Data Quality Management Techniques 2026
Jul 20, 2026 in Guia: Explicação
Master data quality management techniques. Learn to profile, cleanse, & validate data for better decisions & AI readiness.
Não é membro? Registe-se agora
Going beyond default prediction, a continuous system that promotes successful relationships.
Pedro Dias em Jul 27, 2021
Having access to liquidity has been a major issue for humankind for financing both personal life aspects (e.g., housing, cars, college) and for business initiatives (e.g., starting, growing, and expanding a business). The amount of uncertainty that both faces of this exchange – i.e., the creditor and the prospective debtor – face are paramount. On the one hand, the creditor tries to fill in all blanks about the prospect’s capacity to pay the loan requesting additional information, information that in some cases is difficult to obtain by the debtor, resulting in lost opportunities for both sides. On the other hand, the prospect faces a black box interaction, with an approval process that happens behind doors and whose criteria aren’t clear, so the prospective debtor cannot work on adjusting her score.
Artificial Intelligence has played a major role in solving the first side of the equation -the creditor side-, bringing an objective and reproducible credit score estimation of the prospective debtor, at the expense of bringing further opacity to the process for the prospective debtor. In this blog post, we will go through the current paradigm and how we could tackle it differently to perform a more comprehensive approach, identifying current limitations as well as potential solutions from an AI point of view. The ideas described in this article can be applied to both financial entities and subscription-based business models that face default as a source of involuntary churn.
Traditional credit scoring uses blackbox methods -either a human that is out of reach from the end-client or a statistical model- which weights various factors including demographic factors, payment history and other financial indicators. This traditional approach will deny credit to consumers without considering the possibility to adjust their current situation or other extenuating factors. In this sense, it’s modeled as the likelihood at the time of approval of a prospective debtor to pay in full or default. We observe two main limitations:
So, our view on Artificial Intelligence applied to credit scoring is a dynamic interaction, where the model suggests and re-assesses conditions to maximize the mutual benefits on a continuous basis.
So, we can formalize a credit scoring model as a function that, given the prospective debtor, the creditor conditions and the request loan characteristics predicts a valuable KPI for decision making such as the default probability, the expected ROI, etc.
Just for illustrative purposes, such information may comprise:
We can use any model to encode the aforementioned function (e.g., scorecards, decision trees, neural networks).
As mentioned in the previous section, our view on credit scoring considers not just the end decision (i.e., approval vs. rejection) but the adjustment of the loan conditions to ensure a mutually beneficial outcome. So, an AI model should use the aforementioned function to discover the conditions that maximize your KPI
So, we can transform a rejection to an initial submission into approval of an alternative. Please note that while we are simplifying the search for the loan condition adjustment, the model could as well suggest changes to the debtor information. For instance, it could suggest that the easier way of approving the credit is to increase the debtor’s salary by a certain amount, or by reduce his monthly expenses. We prioritize changes on load conditions because they can be applied instantly while looking for approval through the client’s adjustment tend to require longer feedback loops.
However, regardless of our KPI or modeling strategy, we need to be careful about corner cases such as predicting that by lowering the interest rate to zero the chance of a client paying will maximize, leading to a zero profit for the creditor. Similarly, increasing the interest to infinity to maximize profit, but in practice, leads to prospects turning down the loan or getting into bankruptcy. If you’re facing this issue, take a look at causal modeling and monotonic models or simply embed some domain-driven constraints in the decision process.

When dealing with credits, the ability to explain why a certain loan was not conceived, as well as the conditions that were not met, plays a central role. It is not only relevant for those in the receiving end, but also for those who could profit from the deal, if it had happened.
When one models the problem by changing the conditions and context of a client one can infer counterfactual explanations and come up with a hypothesis on how the client could be accepted, and what needs to be changed versus the reality.
Various models have been developed to conceptualize reasons and explanations on the decisions made, an interesting example of this is the model made by Cynthia Rudin’s team for the FICO Data Challenge. The strategy groups features by subscales and attributes to each of these subscales a miniature model (with a softmax). These miniature models are non-linear and thus, by doing so, one can build a tree-like model with machine learning under the hood. A demo of their work can be seen aqui.
From the picture above we can see that the delinquency is a subscale that can be given by those four features and that is contributing to the risk factor (warmer colors represent a higher risk). I would recommend the reader to read Cynthia’s newest article on black box models: “Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead”
Despite all advantages that machine learning models carry, they are also associated with biases. It may be the data you’re using is already biased, the historical decisions made by humans are biased, or the sampling chosen is not representative, or you may be using the wrong algorithm or KPI for the specific problem. The fact is that there will always be some type of bias, and in credit scoring, we have to guarantee that our models are not tending to a specific race, or gender, for example. Note that biased decisions aren’t only prejudicial for the end customer but also for the financial institution. Having models that rely on spurious correlations will eventually lead to catastrophic failure.
Tools like Aequitas, AI Fairness 360, e What-If are open-source toolkits that data scientists can use to check for bias ad discrimination in machine learning models. These tools allow for informed and equitable decisions around developing and deploying predictive tools. More on these tools aqui.
If you want to learn more about Fairness in AI, check our previous post.
A change in the actual paradigm of credit attribution could benefit both parties. More deals could be made, and, more importantly, more deals can be completely fulfilled as time goes by. We live in a fast-paced world, surrounded by continuous changes and new opportunities, why would we oversimplify and assume static solutions to dynamic environments?
Marcar uma reunião com Kelwin Fernandes
Meet Kelwin Saber maisIf you would like to work on credit scoring solutions, or further discuss what was exposed here, feel free to contact us!
Gosta desta história?
Ofertas especiais, últimas notícias e conteúdo de qualidade na sua caixa de entrada.
Jul 20, 2026 in Guia: Explicação
Master data quality management techniques. Learn to profile, cleanse, & validate data for better decisions & AI readiness.
13 de julho de 2026 in Guia: Explicação
Desbloqueie crescimento real com a previsão do valor do tempo de vida do cliente. Aprenda modelos chave, necessidades de dados e roteiros de implementação para resultados estratégicos.
7 de julho de 2026 in Guia: Explicação
Domine a sua estratégia de transformação empresarial com o nosso guia de 2026. Aprenda a executar roteiros de IA, evitar armadilhas e impulsionar o crescimento.
| Bolacha | Duração | Descrição |
|---|---|---|
| cookielawinfo-checkbox-analiticas | 11 meses | Este cookie é definido pelo plugin de Consentimento de Cookies do RGPD. O cookie é usado para armazenar o consentimento do utilizador para os cookies na categoria "Análise". |
| --- O seu texto é uma etiqueta ou nome de campo, provavelmente de um sistema de gestão de cookies ou de um formulário web, e não uma frase completa que necessite de tradução contextual. No entanto, se o objectivo for manter a clareza e a funcionalidade para um utilizador de língua portuguesa, sugiro a seguinte tradução e explicação: **"Checkbox Funcional"** **Explicação:** * **Checkbox:** Refere-se ao elemento gráfico de marcação (uma caixa que pode ser seleccionada ou desmarcada). * **Funcional:** Indica que esta caixa de seleção está relacionada com funcionalidades essenciais do website, como o login, a gestão do carrinho de compras ou outras características que tornam o site utilizável. Se esta etiqueta pertencer a um contexto onde se refere especificamente a cookies, a tradução poderia ser ajustada para ter mais clareza: **"Aceitação de Cookies Funcionais"** ou **"Cookies Essenciais (Funcionais)"** Esta última opção é comum em avisos de cookies para indicar que estes são estritamente necessários para o funcionamento do site. --- | 11 meses | O cookie é definido pelo consentimento de cookies GDPR para registar o consentimento do utilizador para os cookies na categoria "Funcional". |
| cookielawinfo-checkbox-necessary | 11 meses | Este cookie é definido pelo plugin GDPR Cookie Consent. O cookie é usado para armazenar o consentimento do utilizador para os cookies na categoria "Necessário". |
| cookielawinfo-checkbox-outros | 11 meses | Este cookie é definido pelo plugin GDPR Cookie Consent. O cookie é usado para armazenar o consentimento do utilizador para os cookies na categoria "Outros". |
| checkbox-performance-cookielawinfo | 11 meses | Este cookie é definido pelo plugin GDPR Cookie Consent. O cookie é usado para armazenar o consentimento do utilizador para os cookies na categoria "Desempenho". |
| política_de_cookies_visualizada | 11 meses | O cookie é definido pelo plugin GDPR Cookie Consent e é utilizado para armazenar se o utilizador consentiu ou não com a utilização de cookies. Não armazena quaisquer dados pessoais. |