Where Business Models Worked, and Didn’t, and Are Most Needed Now in Mortgages

by Guest Contributor 9 min read February 14, 2012

Part I: Types and Complexity of Models, and Unobservable or Omitted Variables or Relationships

By: John Straka

Since the financial crisis, it’s not unusual to read articles here and there about the “failure of models.” For example, a recent piece in Scientific American critiqued financial model “calibration,” proclaiming in its title, Why Economic Models are Always Wrong. In the mortgage business, for example, it is important to understand where models have continued to work, as well as where they failed, and what this all means for the future of your servicing and origination business.

I also see examples of loose understanding about best practices in relation to the shortcomings of models that do work, and also about the comparative strengths and weaknesses of alternative judgmental decision processes.  With their automation efficiencies, consistency, valuable added insights, and testability for reliability and robustness, statistical business models driven by extensive and growing data remain all around us today, and they are continuing to expand.  So regardless of your views on the values and uses of models, it is important to have a clear view and sound strategies in model usage.

A Categorization: Ten Types of Models

Business models used by financial institutions can be placed in more than ten categories, of course, but here are ten prominent general types of models:

  1. Statistical credit scoring models (typically for default)
  2. Consumer- or borrower-response models
  3. Consumer- or borrower-characteristic prediction models
  4. Loss given default (LGD) and Exposure at default (EAD) models
  5. Optimization tools (these are not models, per se, but mathematical algorithms that often use inputs from models)
  6. Loss forecasting and simulation models and Value-at-risk (VAR) models
  7. Valuation, option pricing, and risk-based pricing models
  8. Profitability forecasting and enterprise-cash-flow projection models
  9. Macroeconomic forecasting models
  10. Financial-risk models that model complex financial instruments and interactions

Types 8, 9 and 10, for example, are often built up from multiple component models, and for this reason and others, these model categories are not mutually exclusive.  Types 1 through 3, for example, can also be built from individual-level data (typical) or group-level data.  No categorical type listing of models is perfect, and this listing is also not intended to be completely exhaustive.

The Strain of Complexity (or Model Ambition)

The principle of Occam’s razor in model building, roughly translated, parallels the business dictum to “keep it simple, stupid.”  Indeed, the general ordering of model types 1 through 10 above (you can quibble on the details) tends to correspond to growing complexity, or growing model ambition.

Model types 1 and 2 typically forecast a rank-ordering, for example, rather than also forecasting a level.  Credit scores and credit scoring typically seek to rank-order consumers in their default, loss, or other likelihoods, without attempting to project the actual level of default rates, for example, across the score distribution.  Scoring models that add the dimension of level prediction increase this layer of complexity.

In addition, model types 1 through 3 are generally unconditional predictors.  They make no attempt to add the dimension of predicting the time path of the dependent variable.  Predicting not just a consumer’s relative likelihood of an event over a future time period as a whole, for example, but also the event’s frequency level and time path of this level each year, quarter, or month, is a more complex and ambitious modeling endeavor.  (This problem is generally approached through continuous or discrete hazard models.)

While generalizations can be hazardous (exceptions can typically be found), it is generally true that, in the events leading up to and surrounding the financial crisis, greater model complexity and ambition was correlated with greater model failure.  For example, at what is perhaps an extreme, Coval, Jurek, and Stafford (2009) have demonstrated how, for model type 10, even slight unexpected changes in default probabilities and correlations had a substantial impact on the expected payoffs and ratings of typical collateralized debt obligations (CDOs) with subprime residential mortgage-backed securities as their underlying assets.  Nonlinear relationships in complex systems can generate extreme unreliability of system predictions.

To a lesser but still significant degree, the mortgage- or housing-related models included or embedded in types 6 through 10 were heavily dependent on home-price projections and risk simulation, which caused significant “expected”-model failures after 2006.  Home-price declines in 2007-2009 reached what had previously only been simulated as extreme and very unlikely stress paths.  Despite this clear problem, given the inescapable large impact of home prices on any mortgage model or decision system (of any kind), it is generally acceptable to separate the failure of the home-price projection from any failure of the relative default and other model relationships built around the possible home-price paths.  In other words, if a model of type 8, for example, predicted the actual profitability and enterprise cash flow quite well given the actual extreme path of home prices, then this model can be reasonably regarded as not having failed as a model per se, despite the clear, but inescapable reliance of the model’s level projections on the uncertain home-price outcomes.

Models of type 1, statistical credit scoring models, generally continued to work well or reasonably well both in the years preceding and during the home-price meltdown and financial crisis.  This is very largely due to these models’ relatively modest objective of simply rank-ordering risks, in general.  To be sure, scoring models in mortgage, and more generally, were strongly impacted by the home price declines and unusual events of the bubble and subsequent recession, with deteriorated strength in risk separation.  This can be seen, for example, in the recent VantageScore® credit score stress-test study, VantageScore® Stress Testing, which shows the lowest risk separation ability in the states with the worst home-price and unemployment outcomes (CA, AZ, FL, NV, MI).  But these kinds of significant but comparatively modest magnitudes of deterioration were neither debilitating nor permanent for these models.   In short, even in mortgage, scoring models generally held up pretty well, even through the crisis—not perfectly, but comparatively better than the more complex level-, system-, and path-prediction models. (see footnote 1)

Scoring models have also relied more exclusively on microeconomic behavioral stabilities, rather than including macroeconomic risk modeling.  Fortunately the microeconomic behavioral patterns have generally been much more stable.  Weak-credit borrowers, for example, have long tended to default at significantly higher rates than strong credit borrowers—they did so preceding, and right through, the financial crisis, even as overall default levels changed dramatically; and they continue to do so today, in both strong and weak housing markets. (see footnote 2)

As a general rule overall, the more complex and ambitious the model, the more complex are the many questions that have to be asked concerning what could go wrong in model risks.  But relative complexity is certainly not the only type of model risk.  Sometimes relative simplicity, otherwise typically desirable, can go in a wrong direction.

Unobservable or Omitted Variables or Relationships

No model can be perfect, for many reasons.  Important determining variables may be unmeasured or unknown.  Similarly, important parameters and relationships may differ significantly across different types of populations, and different time periods.  How many models have been routinely “stress tested” on their robustness in handling different types of borrower populations (where unobserved variables tend to lurk) or different shifts in the mix of borrower sub-populations?  This issue is more or less relevant depending on the business and statistical problem at hand, but overall, modeling practice has tended more often than not to neglect robustness testing (i.e., tests of validity and model power beyond validation samples).

Several related examples from the last decade appeared in models that were used to help evaluate subprime loans.  These models used generic credit scores together with LTV, and perhaps a few other variables (or not), to predict subprime mortgage default risks in the years preceding the market meltdown.  This was a hazardous extension of relatively simple model structures that worked better for prime mortgages (but had also previously been extended there).  Because, for example, the large majority of subprime borrowers had weak credit records, generic credit scores did not help nearly as much to separate risk.  Detailed credit attributes, for example, were needed to help better predict the default risks in subprime.  Many pre-crisis subprime models of this kind were thus simplified but overly so, as they began with important omitted variables.

This was not the only omitted-variables problem in this case, and not the only problem.  Other observable mortgage risk factors were oddly absent in some models.  Unobserved credit risk factors also tend to be correlated with observed risk factors, creating greater volatility and unexplained levels of higher risk in observed higher-credit-risk populations.  Traditional subprime mortgages also focused mainly on poor-credit borrowers who needed cashout refinancing for debt consolidation or some other purpose.  Such borrowers, in shaky financial condition, were more vulnerable to economic shocks, but a debt consolidating cashout mortgage could put them in a better position, with lower total monthly debt payments that were tax deductible.  So far, so good—but an omitted capacity-risk variable was the number of previous cashout refinancings done (which loan brokers were incented to “churn”).  The housing bubble allowed weak-capacity borrowers to sustain themselves through more extracted home equity, until the music stopped.  Rate and fee structures of many subprime loans further heightened capacity risks.  A significant population shift also occurred when subprime mortgage lenders significantly raised their allowed LTVs and added many more shaky purchase-money borrowers last decade; previously targeted affordable-housing programs from the banks and conforming-loan space had instead generally required stronger credit histories and capacity.  Significant shifts like this in any modeled population require very extensive model robustness testing and scrutiny.  But instead, projected subprime-pool losses from the major purchasers of subprime loans, and the ratings agencies, went down in the years just prior to the home-price meltdown, not up (to levels well below those seen in widely available private-label subprime pool losses from 1990’s loans).

Rules and Tradition in Lieu of Sound Modeling

Interestingly, however, these errant subprime models were not models that came into use in lender underwriting and automated underwriting systems for subprime—the front-end suppliers of new loans for private-label subprime mortgage-backed securities.  Unlike the conforming-loan space, where automated underwriting using statistical mortgage credit scoring models grew dramatically in the 1990s, underwriting in subprime, including automated underwriting, remained largely based on traditional rules.

These rules were not bad at rank-ordering the default risks, as traditional classifications of subprime A-, B, C and D loans showed.  However, the rules did not adapt well to changing borrower populations and growing home-price risks either.  Generic credit scores improved for most subprime borrowers last decade as they were buoyed by the general housing boom and economic growth.  As a result, subprime-lender-rated C and D loans largely disappeared and the A- risk classifications grew substantially.

Moreover, in those few cases where statistical credit scoring models were estimated on subprime loans, they identified and separated the risks within subprime much better than the traditional underwriting rules.  (I authored an invited article early last decade, which included a graph, p. 222, that demonstrated this, Journal of Housing Research.)  But statistical credit scoring models were scarcely or never used in most subprime mortgage lending.

In Part II, I’ll discuss where models are most needed now in mortgages.

Footnotes:
[1] While credit scoring models performed better than most others, modelers can certainly do more to improve and learn from the performance declines at the height of the home-price meltdown.  Various approaches have been undertaken to seek such improvements.

[2] Even strategic mortgage defaults, while comprising a relatively larger share of strong-credit borrower defaults, have not significantly changed the traditional rank-ordering, as strategic defaults occur across the credit spectrum (weaker credit histories include borrowers with high income and assets).

Related Posts

What Is AI Decisioning?

Every business makes decisions about people and transactions all day long. Should we approve this loan? Is this purchase fraud? Which customer should get this offer, and what should it be? For a long time, those decisions were made in one of two ways: a person reviewed each case by hand, or the company wrote fixed rules, like "approve anyone with a credit score above 700." Both work. Both also leave value on the table. The manual review is slow and hard to scale. The fixed rule can turn away good applicants and is slow to adapt when the market shifts. AI decisioning is a third way. What makes AI decisioning work Instead of relying on a single reviewer or a rigid rule, automated decisioning uses models that learn from data — studying how thousands of past cases turned out, finding the patterns that predict an outcome, and applying them to each new decision, often in real time. The result is faster, more consistent decisions. But a model on its own isn't the whole story. Getting real value from AI decisioning takes good data to learn from, AI analytics to generate insights, the tools to act on it and the governance to keep it compliant. What we've found is that the pieces only pay off when they work together, and that is where we're built differently. A model is only as good as what it learns from, and we pair your data with one of the deepest views of consumer and commercial credit: decades of full-file history and vetted attributes. Then we give you the tools to act on it. Use cases across your business Whether you're trying to grow your customer base, reduce fraud, manage lending risk, or improve collections, automated decisioning brings all the pieces together to make more accurate, consistent and explainable decisions at scale. Fraud and Identity A fraudulent transaction that slips through costs money and erodes trust. Rules are static, and fraudsters move fast. They'll probe boundaries, find the blind spots and move to the next scheme. By the time the rules are updated, they're already three steps ahead. How AI decisioning changes this: AI fraud detection with real-time risk scoring and decisioning across transactions and customer interactions Intelligence that continuously learns from results to help adapt fraud strategies as threats evolve Reduced false positives and less friction for customers at account opening and checkout Identity verification tools that confirm someone is who they say they are without slowing down the experience Credit and Lending Loan approval is where the relationship begins. Credit risk decisioning helps lenders find that delicate balance between approving enough people to grow, but carefully enough to manage risk. Missing that balance means turning away good customers or taking on losses that are difficult to absorb. How AI decisioning changes this: Increased approval opportunities for creditworthy applicants without increasing overall risk Models you can update and deploy quickly as market conditions change, rather than waiting months Ability to run "what-if" scenarios to test how a new strategy would have performed on your historical data before putting it live Collections Which customer should your team reach out to today? Through which channel? What kind of message? If you reach out too aggressively, you push someone who might have recovered into default. If you wait too long, you lose them. If you call someone at work, they resent you; if you text, they might ignore it. If you offer a payment plan, they might accept it, but only if the terms make sense to their financial situation. How AI decisioning changes this: Optimized next-best-action and contact-channel strategies for each individual customer Improved recovery potential through better targeting Less time spent on accounts with a lower propensity to pay, freeing your team for higher-impact cases Ability to segment and test new strategies before rollout Customer Acqusition Finding the right customers is about reaching the right people with the right offer at the right time. To stay competitive, it’s now a requirement to balance growth with risk while creating a seamless experience converting prospects into customers. How AI decisioning changes this: More precise prospect targeting using credit, behavioral, and alternative data, where permitted, to identify consumers most likely to respond Personalized offers delivered in real time Dynamic decision strategies that can be updated quickly as market conditions and customer behavior change Ongoing testing and optimization of acquisition strategies to improve campaign performance and support customer lifetime value Driving results with AI decisioning Every customer interaction is a decision. Businesses that can adapt quickly will be better positioned to grow, manage risk, and deliver the experiences customers expect. The technology will continue to evolve, but the goal remains the same: making informed decisions that balance business objectives, risk, and customer experience. Learn more about our decisioning software

July 27, 2026 by Zohreen Ismail
Why Innovation Matters for Members First Credit Union

Learn how Members First Credit Union uses innovation and data-driven insights to better serve members and expand financial opportunity.

July 24, 2026 by Scarlet Nickel
Ask the Expert: Unlocking the ROI of alternative data with Natasha Madan and Julius Heim

A visibility gap lenders can't afford to ignore Alternative data is often associated with thin-file or credit invisible consumers. But its value extends far beyond those segments. Experian's Clarity Services database includes approximately one in five credit-active consumers, including one in four consumers with prime-and-above credit profiles. That means lenders may be missing important signals, not only for emerging borrowers, but also for applicants who appear well qualified using traditional bureau data alone. Consider two consumers with the same credit score. Based on traditional credit data, they may appear equally creditworthy. But when Clarity data is added, one consumer may demonstrate stable repayment behavior while another shows recent defaults on alternative finance products. The credit score hasn't changed, but the decisioning context has. That's where alternative data creates value: helping lenders distinguish between consumers who look similar on paper but represent very different levels of risk and opportunity. In this Ask the Expert session, Experian’s Julius Heim, Vice President of Analytics Product Build, Innovation and Scores, and Natasha Madan, Senior Director, Analytics Consulting, explain how different alternative data assets solve different business challenges and why the greatest return comes from using them together throughout the credit lifecycle. What that visibility gap is really costing lenders Better visibility matters because every lending decision carries consequences. Without alternative data, lenders may approve applicants whose repayment behavior suggests elevated risk but isn't reflected in a traditional credit file. Without cash flow insights, they may decline consumers who appear thin file on bureau data despite demonstrating strong income and responsible financial management. The result is a two-sided cost: avoidable bad debt on one side and missed growth opportunities on the other. But ROI extends beyond approvals alone. It also appears through stronger marketing strategies, improved conversion, reduced friction and more precise risk segmentation throughout the lending lifecycle. "ROI can mean many things ... marketing to the right people, achieving better approval rates, reducing risk, getting less friction and overall profitability."Julius Heim, Vice President of Analytics Product Build, Innovation and Scores Where alternative data creates ROI Improve approval strategies Use additional consumer signals to recover creditworthy applicants while avoiding unnecessary declines. Reduce portfolio risk Identify elevated repayment risk earlier through enhanced visibility beyond traditional bureau data. Improve portfolio performance Increase conversion, reduce friction and strengthen profitability across the credit lifecycle. Different data. Different jobs. Not all alternative data solves the same problem. Clarity Services can help lenders strengthen decisions early in the customer journey. It provides additional visibility during prospecting and acquisition, helping identify potential risk before an application moves through the underwriting process. Cash flow insights can provide value in a different way. When traditional credit information offers part of the picture, consumer-permissioned cash flow data can provide greater insight into income, spending patterns and financial capacity. That makes it especially valuable as a second look during underwriting. Together, these complementary data assets help lenders improve decisioning throughout the credit lifecycle. They can support acquisition, underwriting, account management and collections while building on the trusted foundation of traditional bureau data. Research also continues to demonstrate measurable lift when cash flow insights are combined with traditional credit information. "I recently did a study with a client where we actually saw a 20% lift in KS [Kolmogorov-Smirnov] above and beyond credit bureau data. Again, the bureau data itself was very predictive. But even from the cash flow data, we still got a 20% lift, which is an amazing stat." Julius Heim, Vice President of Analytics Product Build, Innovation and Scores The greatest value comes from using these data sources together for a more holistic consumer view. Start with proof, then build Adopting alternative data doesn't have to begin with a large transformation. A practical first step is a data study. By comparing current decision strategies with enhanced data, lenders can identify where additional visibility creates measurable lift within their own portfolios. This approach allows institutions to validate results before making broader operational changes. Every lender has different workflows, technology environments and business priorities. A flexible implementation strategy helps organizations incorporate new data in ways that support existing processes rather than disrupting them. Three ways to get started Run a data study Benchmark current decision strategies and quantify potential lift. Start simple Begin with targeted data attributes or proven scores before expanding to more advanced use cases. Build with confidence Scale implementation based on measured business outcomes and organizational priorities. This approach allows lenders to validate results, build confidence and expand their strategy over time. Explore alternative data with a trusted partner Every lending decision benefits from better consumer insight. Experian helps lenders combine trusted credit data with alternative data, cash flow insights and advanced analytics to strengthen decisioning, improve portfolio performance and uncover new opportunities for growth. Whether you're evaluating alternative data for the first time or expanding an existing strategy, Experian can help you identify where additional consumer insight can create measurable business value. Learn more Contact us About our experts Julius Heim Vice President of Analytics Product Build, Innovation and Scores, Experian Julius Heim works at the intersection of financial services, analytics and innovation. He focuses on leveraging data to drive smarter decision-making and support more inclusive financial ecosystems. Julius brings a practical perspective on how organizations can translate insights into real-world impact, with particular interest in emerging trends across fintech, credit, and the use of alternative data, such as cash-flow data, across the credit lifecycle. Previously, he served as Head of Analytics on the lender side and held roles in insurance analytics earlier in his career. Natasha Madan Senior Director, Analytics Consulting, Experian Natasha Madan partners with lenders to drive smarter, data-driven credit and risk decisions. She specializes in leveraging alternative data and advanced analytics to help organizations improve portfolio performance, optimize customer acquisition, and expand responsible access to credit. During her 15 years at Experian, Natasha has held leadership roles spanning data analytics, product analytics and consulting, giving her a broad perspective of how data can be leverage to solve complex business challenges. She has worked with a diverse range of lenders – including banks, credit unions, fintechs and specialty finance companies to develop analytics strategies that optimize customer acquisition, underwriting and portfolio management. Natasha is passionate about helping organizations unlock the full potential of data to improve both business outcomes and consumer financial inclusion.

July 24, 2026 by Julie.JLee@experian.com