Data Discovery customers, did you know that signing in gives you access to even more features?

Sign in

Analytics Data | Quantitative Analytics

StarMine Text Mining Credit Risk Model

Text mining analytics dataset offering credit risk model signals for quantitative investment and risk assessment.

Key facts for StarMine Text Mining Credit Risk Model

Coverage metric

Geography:
North America, Asia / Pacific, EMEA, Latin America and the Caribbean
History:
From 1998
Coverage count:
Legal entities - 39000 Public companies
Asset Class
Ordinary Shares, Equities

Delivery metadata

Data Frequency
Daily
Language
English
Delivery methods:
Excel, Web Service, API, Deployed/Onsite Servers, SFTP, Cloud, Desktop, FTP, Bulk, Snowflake, Website, RSS Feed
Data formats:
PDF, GZIP, XML, JSON, SQL, CSV, Text, Python, MPEG, HTML, Bitmap, User Interface, PCAP
Minimum service frequency
Daily

Overview of StarMine Text Mining Credit Risk Model

Description of the dataset

• LSEG provides the StarMine Text Mining Credit Risk Model to assess corporate financial distress risk using language signals from trusted textual sources. We evaluate content from Reuters News, StreetEvents conference call transcripts, corporate filings and selected broker research reports.

• Our model applies a bag-of-words text mining methodology that analyses the frequency of words and phrases linked to observed credit outcomes. We transform unstructured language into a systematic credit health signal designed to identify companies more likely to weaken or thrive.

• LSEG ranks publicly traded companies on a 1-100 percentile scale, with 100 representing the healthiest companies. We support credit surveillance, equity risk assessment and early-warning workflows by incorporating information that may not be fully captured in traditional financial statement data.

Accessing the dataset

This dataset can be used by the following products. Talk to us to learn more about different packages and offerings.

LDMS - LSEG Data Management Solution

Cloud-hosted data management and repository platform for commodities content. Supports loading and consolidation of LSEG, third-party, and proprietary data and delivers near-streaming and non-streaming datasets such as pricing, reference data, and fundamentals via a unified interface for integration and analysis.