Bloomberg USA Daily News Headlines & Articles Dataset
A daily file of Bloomberg's US-edition coverage for your keyword list — 16 columns covering headline, standfirst, article text, journalists, keyword tag, timestamp to the second, permalink, Bloomberg's article ID and images.
The sample's keywords track Gulf oil executives and assets — Aramco's Amin Nasser under three spellings, Iraq's national oil company, Kuwait's air bases — and captured 50 articles across two days, from India's capital flows to the yen and Hong Kong tech.
The file already exists, because this pipeline runs daily whether you buy it or not. Most vendors start collecting after you order — which is why they quote a lead time. Here the most recent file is in your account within minutes of payment, with an API key issued at the same time.
What you get, in plain terms
Four things. The once-a-day version of our Bloomberg feed — a complete daily record for your keywords.
The keyword that caught it
Keyword records which tracking term captured each article, so every row arrives already tagged to the company, person, place or theme you are watching.
Name variants tracked together
One entity, several spellings: “Amin H Naser”, “Amin Hasan Nasser” and “Amin Nasser said” all route to the same person. The keyword list can hold as many variants as you need.
Articles and video summaries
Written stories carry their text; Bloomberg TV interviews carry a summary. The permalink and standfirst tell you which is which.
However you want it
Direct download, REST API, Amazon S3, Google Cloud, Snowflake or SFTP. The API key comes with the dataset.
Fields included in this dataset
All 16 columns, in the exact order they appear in the file — taken straight from the sample, not from a brochure. The free sample ships with a data dictionary giving an example value for each one.
Sample rows from the real file
Real rows from the sample file, not an illustration. The supplied sample has 50 articles with all 16 columns.
Coverage
Captured daily across bloomberg.com's US-edition news and opinion pages for the keyword list. One row is one article, identified by Bloomberg's Article ID and permalink.
| Metric | In the supplied sample | What it tells you |
|---|---|---|
| Articles | 50 across 44 hours | 9 to 11 August 2026 |
| Distinct keywords | 29 | Executives, oil companies, air bases, places |
| News · Opinion | 47 · 3 | From the permalink section |
| Short video or audio summaries | 13 of 50 | Under 60 words — TV interviews and podcasts |
| Journalist named | 32 of 50 | “Bloomberg”, “Bloomberg News” or blank on the rest |
Reach · Web Shares |
0 of 50 | Not published by Bloomberg — see the FAQ |
Need stories within hours? A four-times-a-day Bloomberg feed is also available. Or send your own keyword list and we will run the feed against it.
Historical data
The last 30 days come with the dataset. Beyond that we hold Bloomberg USA records from February 2025 onwards.
Ask about historical dataNeed more data points?
We can extend this dataset beyond the standard 16 columns — a content-type flag separating articles from video is the most requested addition, along with entity tags for companies and tickers.
Request custom fields“We follow a handful of Gulf executives. The daily file catches every Bloomberg mention, under every spelling, and lands before our morning meeting.”
What people use this dataset for
Media monitoring and PR
Track every mention of your company, executives and competitors the day it is published.
Investment and risk teams
Feed timestamped headlines and article text into event studies, sentiment and alerting models.
Market and policy intelligence
Follow a sector, a country or a regulator through the outlet its decision-makers read.
Data and AI teams
Build search, retrieval and evaluation sets on clean, structured article text with full provenance.
About Bloomberg USA news data
Bloomberg's US edition is the version of bloomberg.com most global finance teams read. This daily feed captures its public article pages against a keyword list, giving you one structured file a day of every Bloomberg story that touches the entities you track.
What the sample was tracking
The keyword list followed Gulf energy: Aramco's chief executive under several spellings, Iraq's state oil company, the Abu Safah field and Kuwait's Ahmad Al Jaber Air Base. The results were broad — a former RBI governor on capital leaving India, the Bank of Korea signalling more hikes, Amkor exploring a stake sale in its China unit, Hong Kong's tech index overhaul — because Bloomberg mentions energy names across its whole report.
Articles and video in one file
13 of 50 sampled rows were Bloomberg TV or audio items with a summary under 60 words, such as a Titan CFO interview or “Why is the yen losing gains from intervention?”. Filter on length or the standfirst if you only want written stories.
Two things to plan for
Keyword appeared in the captured text on only 6 of 50 rows — it usually sits deeper in the full story — so treat it as a routing tag. Two headlines appear twice under different permalinks where Bloomberg republished an updated story; de-duplicate on Article ID.
Is collecting this data legal?
We collect only publicly accessible pages, respect robots.txt and platform terms, and align with GDPR and CCPA — the only personal data in the file is the public byline a publisher prints. Article text remains the publisher's copyright: the standard licence covers internal analysis such as monitoring, research and signal extraction, not republishing. Each row carries its permalink and publication time as documented provenance, so your legal team can review the source before you buy.
Frequently asked questions
Keyword records the tracking term that captured each article. Send us your own list — companies, people, places, tickers — and we will run the feed against it at no extra cost.
