Machine Learning for Real-World Credit Risk

by Alan Ikemura 3 min read September 12, 2018

Machine learning (ML), the newest buzzword, has swept into the lexicon and captured the interest of us all. Its recent, widespread popularity has stemmed mainly from the consumer perspective. Whether it’s virtual assistants, self-driving cars or romantic matchmaking, ML has rapidly positioned itself into the mainstream.

Though ML may appear to be a new technology, its use in commercial applications has been around for some time. In fact, many of the data scientists and statisticians at Experian are considered pioneers in the field of ML, going back decades. Our team has developed numerous products and processes leveraging ML, from our world-class consumer fraud and ID protection to producing credit data products like our Trended 3DTM attributes. In fact, we were just highlighted in the Wall Street Journal for how we’re using machine learning to improve our internal IT performance.

ML’s ability to consume vast amounts of data to uncover patterns and deliver results that are not humanly possible otherwise is what makes it unique and applicable to so many fields. This predictive power has now sparked interest in the credit risk industry. Unlike fraud detection, where ML is well-established and used extensively, credit risk modeling has until recently taken a cautionary approach to adopting newer ML algorithms. Because of regulatory scrutiny and perceived lack of transparency, ML hasn’t experienced the broad acceptance as some of credit risk modeling’s more utilized applications.

When it comes to credit risk models, delivering the most predictive score is not the only consideration for a model’s viability. Modelers must be able to explain and detail the model’s logic, or its “thought process,” for calculating the final score. This means taking steps to ensure the model’s compliance with the Equal Credit Opportunity Act, which forbids discriminatory lending practices. Federal laws also require adverse action responses to be sent by the lender if a consumer’s credit application has been declined. This requires the model must be able to highlight the top reasons for a less than optimal score.

And so, while ML may be able to deliver the best predictive accuracy, its ability to explain how the results are generated has always been a concern. ML has been stigmatized as a “black box,” where data mysteriously gets transformed into the final predictions without a clear explanation of how. However, this is changing.

Depending on the ML algorithm applied to credit risk modeling, we’ve found risk models can offer the same transparency as more traditional methods such as logistic regression. For example, gradient boosting machines (GBMs) are designed as a predictive model built from a sequence of several decision tree submodels. The very nature of GBMs’ decision tree design allows statisticians to explain the logic behind the model’s predictive behavior. We believe model governance teams and regulators in the United States may become comfortable with this approach more quickly than with deep learning or neural network algorithms. Since GBMs are represented as sets of decision trees that can be explained, while neural networks are represented as long sets of cryptic numbers that are much harder to document, manage and understand.

In future blog posts, we’ll discuss the GBM algorithm in more detail and how we’re using its predictability and transparency to maximize credit risk decisioning for our clients.

Related Posts

Ask the Expert: The Future of Lending Starts With Identity With Shawn Rife and Brian Cardona

Identity intelligence and alternative data can help lenders validate consumers and support more informed decisions across the customer lifecycle.

September 16, 2026 by Julie Lee
Financial Institutions Are Rethinking Customer Acqusition

Customer acquisition strategies are constantly evolving toward more precise targeting. From a marketing lens, you can track every step, optimize communication channels and still miss the person most likely to convert. Attribution can tell us which channels work and automation can make marketing spend more efficient. But both assume we know who is actually on the other end. Financial institutions are learning that finding audiences and targeting them is no longer the biggest challenge. As acquisition optimization marketing becomes more sophisticated, teams can measure and act on more signals than before. What they can't always know is whether the person on the receiving end is real. Customer acquisition has evolved into an identity problem. The challenge is not that every questionable signal represents malicious activity. It's that acquisition systems must make increasingly intelligent decisions with an imperfect understanding of who they're actually engaging. When identities are fragmented, duplicated, temporary or synthetic, optimization becomes a question of trust as much as targeting. When your signals don't reliably identify customers The customer journey often includes searching, filling out a form, creating an account, requesting a quote and subscribing. All of these signals work well when identity is relatively stable.  However, financial institutions are finding that these signals are becoming less reliable. A single person can operate across multiple personas, devices, browsers, aliases, accounts and intermediaries while several apparent “people” may actually represent one underlying actor. Financial instituions are finding: Fragmented customer signals Difficulty distinguishing an old account from a new one Different digital pathways associated with the same individual Signals that are generated by automation Real customers getting flagged because signals are too thin to evaluate confidently Legacy signals continue to be challenged Marketing has historically treated intent as a valuable signal because intent was relatively difficult to produce. A search required human intent. A form required someone to fill it out. An inquiry implied a meaningful amount of human effort. Financial institutions are already combating AI-enabled fraud, and now marketing teams are starting to face it on a massive scale. AI can mimic human behavior by researching products, comparing prices, filling out forms, creating accounts and signing up for services. A valid email address is no longer enough. Marketers need to know: How long has it existed? How recently has it been active? Does its activity appear consistent or suddenly anomalous? Has it gone dormant and returned? Is it associated with patterns that suggest stability or unusual behavior? How to build on your strongest signal Email remains one of the most persistent identifiers in digital commerce, following people across devices, platforms, transactions, subscriptions, accounts and years of activity. For over two decades, this has shaped how AtData thinks about identity. Now, as part of Experian, it’s shaping how an entire platform and team approach identity. A marketer doesn’t need every prospect to have existed online for twenty years. But understanding whether a newly acquired prospect has meaningful identity context can dramatically improve the quality of the decision being made around it. Better identity intelligence can help organizations reduce unnecessary friction by improving their ability to recognize legitimate customers. With a strong identity foundation, marketing teams can better address: Which audiences are more likely to convert? Which leads are high quality? Which channels are driving incremental growth? What do the best prospects look like? The value isn't simply having an email address. It's understanding the history and behavioral context associated with it. That context can provide a stronger digital identity signal, helping marketers understand how long they have been active, whether its behavior is consistent with that of a real person and whether current activity aligns with past patterns. It continues to be one of the most persistent identifiers in digital commerce. An infrastructure built for what's coming The acquisition of AtData by Experian reflects a fundamental shift in how identity infrastructure needs to work. Experian's scale and decisioning capabilities, combined with AtData's real-time email intelligence, create a strong platform. Read more about the why behind the acquisition and see how email works as an identity anchor for fraud prevention. Contact us to learn about our customer acquisition solutions

September 15, 2026 by Zohreen Ismail
As Electric Vehicle Adoption Eases, Dealers Can Find New Opportunities To Reach Consumers

After years of rapid growth, new electric vehicle (EV) registrations have moderated, and the EV market has entered a new chapter. But slower growth shouldn’t be mistaken for disappearing demand, with data suggesting the reality is much more nuanced. According to Experian Automotive’s Automotive Consumer Trends Report: Q2 2026, battery EVs accounted for 8.21% of new retail registrations in the last 12 months, down from 9.23% a year earlier. However, consumers aren’t simply walking away from electrification. In fact, more than one million new EVs were registered during the past 12 months and the used EV market recorded more than 540,000 registrations over the same period. The opportunity may be less about waiting for the EV market to grow and more about understanding where EV demand is present, who is driving them, and how to reach those consumers more effectively. Who is likely to purchase an EV and what vehicle types are they interested in? Understanding who’s in the market for an EV can allow dealers to position themselves around consumers’ needs as they choose a vehicle that fits their everyday lifestyle. In the second quarter of 2026, Millennials and Gen X accounted for 67.83% of new EV registrations, nearly 10 percentage points above their combined share of all new, retail registrations. Millennials were also the largest generational audience across both new and used EV market share, coming in at 35.76% and 38.42%, respectively. It’s important to consider that the EV shopper isn’t necessarily looking for an unfamiliar or new type of vehicle. In many cases, they’re seemingly looking for an electric version of the practical vehicle they already know. For instance, SUVs accounted for 77.47% of new EV registrations in Q2 2026, which was similar to SUVs’ 63.49% share of all new retail registrations. For these shoppers, creating messaging around value, practicality, and available choices may resonate differently than premium technology messaging aimed at some new-EV prospects. The more precisely dealers can identify those audiences, the less they need to depend on broad EV market momentum to generate demand. To learn more about EV insights, view the full Automotive Consumer Trends Report: Q2 2026 presentation.

September 15, 2026 by Kirsten Von Busch

Subscribe to our Newsletter

Enter your name and email for the latest updates.

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Subscribe to our Newsletter

Don't miss out on the latest industry trends and insights!
Subscribe