Live API Benchmark · 2026API 实测对比 · 2026

Best Financial News API
for AI Agents
面向 AI Agent 的
最佳金融新闻 API

A live 2026 benchmark of seven financial-news and retrieval options across ticker lookup, language coverage, relevance, latency, sentiment, and structured fields.

基于实时调用,对七种金融新闻与检索方案的股票代码查询、语言覆盖、相关度、延迟、情绪字段和结构化能力进行对比。

By QVeris ResearchTests run July 30–31, 2026Updated July 31, 2026
QVeris Research测试时间:2026 年 7 月 30–31 日更新:2026 年 7 月 31 日

Short answer: the best financial news API depends on the agent workflow. Across 406 publishable live observations, Finlight was the strongest structured ticker news API (20/20 valid calls, relevant results in both rounds for 7 of 9 listing markets), Linkup had the broadest observed multilingual retrieval (a Web retrieval product, not a strict news feed), EODHD was the best broad-exchange ticker alternative, and Gildata and Caidazi led the Chinese-market workflows.

Key findings at a glance

If your AI agent needs… Use Evidence from the live run
Ticker-first structured news Finlight Financial News 7/9 markets passed both rounds; 20/20 valid calls; clean invalid-ticker control
Broad exchange-coded ticker coverage EODHD Financial News 7/9 markets passed both rounds; provider errors on JP/IN
Global company & macro discovery Linkup Search Relevant results in 18/18 company and 12/12 macro observations
China / HK company news Gildata Stock News, EODHD Relevant results in 4/4 applicable observations
Chinese specialist feeds Caidazi Hybrid Search V2, Gildata Public Opinion 6/6 relevant rounds, 100% top-5 relevance
Familiar English keyword search NewsAPI, Brave News Search 8/10 and 5/6 relevant macro rounds

There is no honest single winner across global discovery, strict ticker filtering, multilingual retrieval, sentiment, and regional specialist feeds. This guide keeps those product shapes separate.

What we tested: 7 financial news APIs, 406 live observations

The broad benchmark completed 406 publishable observations from 410 live API invocations across 203 tool-case cells. Every cell ran twice. Cases covered company, macro, crypto, forex, invalid input, and relevant specialist feeds across nine market regions and seven requested languages.

The ticker benchmark added 120 completed observations across nine listing markets and an invalid ticker control. Two suppliers met the multi-market ranking gate. A market counted as supported only when both rounds returned news relevant to the target company. Each supplier received its own native ticker format rather than one shared symbol dialect.

All overview visuals use one fixed seven-supplier public comparison shortlist: Brave, Caidazi, EODHD, Finlight, Gildata, Linkup, and NewsAPI. The shortlist is the union of suppliers that qualified in at least one published buyer scenario in the baseline edition, including the strict ticker workflow. It is now frozen for comparable retests: a shortlisted supplier that later performs poorly becomes Not qualified but remains one of the seven rows. This also keeps a specialist such as Finlight visible in the global-company view even when that broad input does not clear the scenario gate.

Global company discovery
Relevant-result repeatability in applicable live rounds
Qualification also depends on the published gate and sample scope.
Linkup
100% (18)
Gildata
89% (18)
EODHD
78% (18)
Caidazi
56% (18)
NewsAPI
38% (16)
Finlight
11% (18)
Brave
100% (4)

Brave's four applicable rounds were below the scenario's qualification scope. Percentages are dated observations, not universal provider scores.

“Valid API response,” “scenario returned the expected empty/non-empty shape,” and “the returned news was relevant” are different metrics. Treating them as one success rate would reward irrelevant fallback results and punish honest empty responses.

To prevent a global-search chart from hiding regional or specialist strengths, we compare five buyer scenarios. Every overview row uses the same seven shortlisted suppliers. Qualified means the supplier's best applicable tool cleared that scenario's published gate; Not qualified means it was tested but fell below the gate; N/A means no tested tool from that supplier exposed the applicable contract. A percentage is the share of applicable live rounds that returned at least one relevant result, with the observation count in parentheses.

Each scenario's detailed leaderboard remains narrower: it includes only qualified suppliers and only one tool per supplier. When several tools from the same supplier qualify, the representative is selected by relevant-result repeatability, then top-five relevance, applicable-case count, expected-response rate, valid-response rate, field completeness, and latency.

Workflow fit map
Qualification depends on the job the agent must perform
QualifiedNot qualifiedN/A
SupplierGlobalChina/HKSpecialistMacroTicker
BraveNQQ
CaidaziQQQQ
EODHDQQQ
FinlightNQNQNQNQQ
GildataQQQQ
LinkupQQQ
NewsAPINQNQQ
Supplier Global company discovery China/HK company news China specialist news Macro/cross-asset Multi-market ticker
Brave Not qualified — Brave News Search, 100% (4) N/A N/A Brave News Search — Qualified, 83% (6) N/A
Caidazi Caidazi Hybrid Search V2 — Qualified, 56% (18) Caidazi Hybrid Search V2 — Qualified, 50% (4) Caidazi Hybrid Search V2 — Qualified, 100% (6) Caidazi Hybrid Search — Qualified, 50% (12) N/A
EODHD EODHD Financial News — Qualified, 78% (18) EODHD Financial News — Qualified, 100% (4) N/A N/A EODHD Financial News — Qualified, 7/9 markets
Finlight Not qualified — Finlight Financial News, 11% (18) Not qualified — Finlight Financial News, 0% (4) Not qualified — Finlight Financial News, 0% (6) Not qualified — Finlight Financial News, 33% (12) Finlight Financial News — Qualified, 7/9 markets
Gildata Gildata Public Opinion — Qualified, 89% (18) Gildata Stock News — Qualified, 100% (4) Gildata Public Opinion — Qualified, 100% (6) Gildata Public Opinion — Qualified, 75% (12) N/A
Linkup Linkup Search — Qualified, 100% (18) Linkup Search — Qualified, 100% (4) N/A Linkup Search — Qualified, 100% (12) N/A
NewsAPI Not qualified — NewsAPI Everything, 38% (16) Not qualified — NewsAPI Everything, 0% (4) N/A NewsAPI Everything — Qualified, 80% (10) N/A

This is a qualification map, not a universal quality score. Below-gate entries for shortlisted suppliers remain visible as Not qualified, while detailed leaderboards include qualified entries only. The detailed metrics below distinguish relevance, freshness, fields, source diversity, duplicate behavior, sentiment coverage, latency, and negative controls.

The matrix uses each product's broad discovery or feed input. The strict ticker section is a separate suite using dedicated native ticker parameters and provider-specific symbol formats. Its results are not interchangeable with broad keyword results.

How we scored the financial news APIs: the published evaluation standard

The scoring unit is one observation: one product, one applicable test case, and one live round. A cell is one product-case pair, and an API invocation is one physical request. We report all three to keep the scoring denominator separate from the traffic volume.

Metric Public calculation and pass rule
Valid API response rate Observations with a valid success envelope divided by completed observations. A valid empty result counts as an API success; a provider error does not.
Expected-response rate Positive cases pass when at least one article is returned. The invalid-input control passes only when no article is returned.
Relevant-result round rate Applicable live rounds returning at least one relevant top-five row divided by completed applicable rounds.
Top-5 relevance Relevant rows among the first five results. A row is relevant when the requested entity or topic alias appears in its title, summary, or structured entity fields.
Fresh-record rate Returned rows with a usable publication time inside the case's seven-day window divided by all returned rows; a missing time does not receive freshness credit.
Core-field completeness Present values for title, URL, publication time, source, and summary divided by the expected core-field values in returned rows.
Observable source diversity Distinct normalized source values per non-empty observation. Zero means no usable normalized source field, not zero upstream publishers.
Duplicate rate Repeated normalized URLs or title-and-date identities divided by returned rows.
Sentiment coverage Relevant rows with an exposed sentiment value divided by relevant rows. This measures availability, not sentiment accuracy.
Requested-language rate Rows matching the requested language divided by rows whose language is explicit or detectable. The observation denominator is always shown.
Requested-country rate Rows matching the requested country divided by rows with observable country metadata in cases requesting a concrete country. The observation denominator is always shown.
Median latency Median end-to-end elapsed time across completed live observations. It is an observation from this run, not an SLA.
Strict ticker-market pass A real market passes only when both rounds return news relevant to the target company through the provider's dedicated ticker or symbol input.
Negative-control pass The invalid query or ticker returns no article. An irrelevant fallback is a failure, not a successful result.

At the individual case level and in the supplier comparison matrix, N/A means the inspected endpoint contract cannot express that request. Not qualified means applicable cases were completed but the best supplier tool did not clear the scenario gate. We do not convert contract-level N/A into a failure, and we do not grant untested credit based on another endpoint from the same supplier. Provider errors, wrong-language output, stale or irrelevant fallback data, and unexpected empty results remain in the underlying evaluation. Every comparison must complete two live rounds for every applicable cell before it can be considered for ranking.

A scenario ranking requires at least 50% valid API responses, relevant results in at least 50% of applicable rounds, and an endpoint contract covering at least 50% of that scenario's fixed cases. Only the best qualifying tool from each supplier is shown. The multi-market ticker ranking separately requires at least 50% valid calls and repeatable relevant results in at least five of the nine tested markets. These are eligibility gates, not extra points added after the test.

Ticker syntax is normalized per provider before the final run. Language claims are split into three separate facts: what the vendor documents, what the current integration exposes, and what the fixed live cases actually returned. Charts are generated from the same sanitized scorecard as the tables.

Is this evaluation scientific and accurate?

It is a controlled, reproducible product-selection benchmark, but it is not a statistically conclusive academic study or a long-term reliability certification. The fixed cases, explicit denominators, two-round completion gate, provider-native ticker formats, negative controls, and preserved run evidence make comparisons auditable and useful for choosing an API.

The main limitation is sample size and time dependence. Two rounds can confirm repeatability during the test window, but cannot estimate a provider's monthly SLA, full language inventory, or long-run recall with narrow confidence intervals. News availability also changes with the news cycle, licensing, subscription tier, and indexing delay. We therefore publish observed results with their test date and denominators, avoid a single universal score, and recommend quarterly reruns plus reruns after material integration changes.

Best financial news API by agent workflow

Agent workflow Best observed fit Why
Structured multi-market ticker news Finlight Financial News 20/20 valid calls; relevant results in both rounds for 7/9 markets; clean invalid-ticker behavior
Global company discovery Linkup Search Relevant results in 18/18 observations across all nine fixed company cases
China/HK company news Gildata Stock News, EODHD, and Linkup Each returned relevant results in 4/4 applicable China/HK observations
Chinese specialist news Caidazi Hybrid Search V2 and Gildata Public Opinion Both returned relevant results in 6/6 observations with 100% top-5 relevance
Macro and cross-asset discovery Linkup Search Relevant results in 12/12 observations; NewsAPI followed at 8/10
English discovery Brave News Search Relevant results in 5/6 macro and cross-asset observations within its tested scope
Multi-market ticker alternative EODHD Financial News Relevant results repeated in 7/9 markets

China/HK company-news trade-offs

The table below shows one qualifying representative tool per supplier. All four tools cover both fixed China/HK company cases and returned relevant results in at least half of their live rounds.

Product Relevant rounds Top-5 relevance Fresh records Core fields Sources/non-empty run Duplicate rate Sentiment on relevant rows Median latency Invalid control
Caidazi Hybrid Search V2 2/4 100% 20% 100% 2.0 0% 0% 1,809 ms 0/2
EODHD Financial News 4/4 100% 0% 60% 0.0 0% 100% 2,194 ms 0/2
Gildata Stock News 4/4 100% 0% 60% 0.0 0% 0% 11,056 ms 2/2
Linkup Search 4/4 55% 0% 40% 0.0 0% 0% 4,042 ms 0/2

These numbers expose different product shapes. EODHD and Gildata Stock News were highly relevant in these fixed company cases, while Caidazi exposed richer observable source and core-field data. A 0.0 source count means the normalized response did not expose a usable source value; it does not mean that the supplier ingested zero publishers.

Chinese specialist-news trade-offs

This cohort requires coverage of at least two of the three fixed fund, industry, and organization cases, then keeps only the strongest qualifying tool from each supplier.

Product Observations Relevant rounds Top-5 relevance Fresh records Core fields Sources/non-empty run Duplicate rate Median latency
Caidazi Hybrid Search V2 6 6/6 100% 33% 100% 4.0 0% 2,028 ms
Gildata Public Opinion 6 6/6 100% 53% 48% 0.0 0% 24,739 ms

Both representatives returned relevant results in all six observations. Their trade-off is different: Caidazi was faster and exposed more complete normalized fields and sources, while Gildata returned a higher share of records inside the freshness window.

Macro and cross-asset trade-offs

The following supplier representatives met all ranking gates in the macro, crypto, and forex scenario. Applicability still differs: NewsAPI had 10 observations, while the other rows had 6 or 12.

Product Observations Relevant rounds Top-5 relevance Fresh records Core fields Sources/non-empty run Duplicate rate Median latency Invalid control
Brave News Search 6 5/6 96% 0% 60% 0.0 0% 1,991 ms 0/2
Caidazi Hybrid Search 12 6/12 32% 44% 73% 4.2 8% 2,902 ms 0/2
Gildata Public Opinion 12 9/12 89% 0% 60% 0.0 0% 15,909 ms 2/2
Linkup Search 12 12/12 100% 0% 40% 0.0 0% 6,372 ms 0/2
NewsAPI Everything 10 8/10 70% 35% 100% 2.6 0% 2,826 ms 2/2

Freshness and field values are measured only from normalized returned rows. An invalid-control failure means the endpoint returned fallback content for the deliberately nonexistent query; it does not invalidate its positive-case results, but agents should add their own relevance guard.

Finlight: best ticker news API for structured agent workflows

Finlight completed 20/20 calls in the strict ticker suite, passed both invalid-ticker controls, and returned relevant news in both rounds for US, HK, JP, DE, FR, IN, and ES. CN and BR returned empty results in both rounds. Median latency was 1,574 ms.

Observed ticker workflow
Repeatable relevance across nine listing markets
Passed both roundsEmpty or provider error
Finlight
USHKJPDEFRINESCNBR
7 / 9
EODHD
USHKCNDEFRBRESJPIN
7 / 9

A “pass” means relevant results appeared in both fixed live rounds. It is an observed benchmark result, not a permanent coverage guarantee.

Finlight also returned structured article fields useful to agents: title, URL, publication time, source, summary, language, countries, sentiment label, categories, and expanded company entities. Its confidence field is classification confidence, not directional sentiment, so we did not score it as a sentiment value.

Which languages do financial news APIs support?

Language coverage has three layers:

  1. what the vendor documents;
  2. what the current integration lets an agent request;
  3. what returned relevant requested-language news in live tests.

The chart preserves the same seven-supplier public comparison shortlist. Suppliers for which we did not independently verify a vendor or integration language-control profile remain visible as N/A; they are not treated as supporting zero languages. The compact table lists the four verified profiles and reports observed case results without inferring an undocumented language inventory.

Product Documented or exposed language control Repeatably observed requested languages
Finlight Financial News Vendor API accepts one ISO 639-1 language per request and defaults to en; current integration does not expose the parameter en
Linkup Search No explicit language filter in the tested contract de, en, es, ja, pt, zh; the fr case met the strict rule in 1/2 live rounds
NewsAPI Everything 14 exposed codes: ar, de, en, es, fr, he, it, nl, no, pt, ru, sv, ud, zh en
Brave News Search Current integration enumerates ar, bg, bn, ca, en, eu en
Observed language workflow
Requested-language retrieval is not the same as documented support
NewsAPI

14 language codes exposed

Observed: EN
Documented options did not equal repeatable multilingual retrieval in the fixed company cases.
Finlight integration

Structured ticker strength

Observed: EN
The tested integration exposed English, while native provider documentation describes broader controls.

Finlight's own documentation says that the API indexes news in many languages, filters one language per request, and defaults to English. It also documents strict ticker filtering and exchange-aware company entities. The missing language control is therefore an integration gap, not evidence that the underlying Finlight API is English-only. See the Finlight REST endpoint and multilingual query guide.

NewsAPI documents its exact language enum on the Everything endpoint.

Stock ticker lookup by market: which API resolves symbols reliably?

Product Markets passing both rounds Edge-case behavior
Finlight Financial News US, HK, JP, DE, FR, IN, ES Invalid ticker returned empty in both rounds
EODHD Financial News US, HK, CN, DE, FR, BR, ES JP, IN, and invalid ticker returned provider errors

Ticker support is not the same as keyword recall. A product can find an article containing “Toyota” yet fail a strict 7203.T lookup, and different providers use different suffixes for the same listing. Agents that accept user-entered symbols should normalize symbols per provider before calling the news endpoint.

Country and requested-language results

This matrix preserves the same seven-supplier public comparison shortlist and uses each supplier's global-company representative, whether qualified or below the gate. Every applicable tool-case cell was called twice; N/A means the case is outside that endpoint's tested contract, not that the test was skipped.

Each percentage is the number of passing live rounds divided by two. A round passes only when the call has the expected response shape, returns at least one relevant result, all language-observable top-five results match the requested language, and any available country metadata matches the requested market. Therefore, 50% means one of two live rounds passed the complete rule; it is not an article-level language accuracy or a long-term provider support rate.

Strict case rule
A market-language cell passed only when every check passed
2 live rounds per cell
Expected response shape+Relevant result+Requested language+Country metadata, when available=PASS

A 50% cell means one of two complete live rounds passed. It is not an article-level language-accuracy score.

Country tags in financial news can mean source country, company domicile, listing market, or article topic. Buyers should verify which semantic a provider uses before treating country filters as exchange filters.

How to choose a financial news API for your AI agent

Choose Finlight when your agent starts with a ticker and needs structured entities, sentiment labels, categories, provenance, and clean negative-control behavior. If multilingual Finlight retrieval matters, expose its native language parameter first and rerun the same fixed cases.

Choose EODHD when broad exchange-coded ticker coverage matters and your agent can handle provider errors explicitly.

Choose Linkup when multilingual recall and citations matter more than having a strict news-feed contract. Choose Brave for English discovery within its tested scope. Choose NewsAPI when you need a familiar keyword API and explicit language selection, while accepting that declared language options do not guarantee relevant local-market results.

For China/HK company workflows, compare Gildata Stock News, EODHD, Linkup, and Caidazi Hybrid Search V2 against the response shape your agent needs. Gildata Stock News and EODHD were the most consistently relevant in the two fixed company cases; Caidazi exposed more complete normalized fields and observable source diversity; Linkup traded structure for broader retrieval.

For Chinese fund, industry, and organization workflows, Caidazi Hybrid Search V2 was faster and more structurally complete in the fixed suite. Gildata Public Opinion matched all six relevant-result rounds with stronger observed freshness, but higher latency.

For macro, crypto, and forex discovery, Linkup had the strongest relevant-result repeatability, followed by NewsAPI and Gildata Public Opinion in their applicable cases. Linkup and Brave both failed the invalid-query control, so an agent using them should validate entity/topic relevance before accepting fallback results.

Limitations

This is a dated live observation, not a permanent statement about a provider's entire inventory. News availability changes with the news cycle, subscription tier, source licensing, indexing delay, and integration schema. Empty results do not prove a market is contractually unsupported; they mean that the fixed query returned no relevant article in the tested window.

The benchmark scores the endpoints and parameters actually called. It does not transfer a capability from another endpoint owned by the same provider.

How providers can join or improve the benchmark

Providers can submit a comparable news discovery or ticker-news endpoint, document its market and language controls, and supply a test credential or QVeris integration. Inclusion and ranking cannot be purchased. Corrections to contract facts can be made from documentation; performance results change only after the fixed live suite is rerun.

Start by checking whether your product is already listed in the QVeris Provider Hub. To request a new listing, correct an endpoint contract, or schedule a comparable retest, email support@qveris.ai with “Financial News Benchmark” in the subject.

For the fastest retest, provide:

  • the exact endpoint and stable schema;
  • supported ticker formats and exchanges;
  • language codes and whether the filter is article-level or source-level;
  • country-field semantics;
  • source, licensing, and freshness constraints;
  • invalid-input and empty-result behavior.

FAQ: financial news APIs for AI agents

What is the best financial news API for an AI agent?

There is no single winner. In this benchmark, Finlight was the best structured ticker news API, Linkup the strongest for multilingual and macro discovery, EODHD the best broad-exchange ticker alternative, and Gildata and Caidazi the best for Chinese-market coverage. Pick by workflow, not by an overall score.

Can I search financial news by stock ticker?

Yes. Several tested integrations exposed dedicated ticker or symbol inputs, but only Finlight and EODHD met the published multi-market ranking gate. Market coverage and ticker syntax differ, so normalize the user's symbol to the provider's dialect (for example 7203.T versus TYO:7203).

Which API had the broadest observed language coverage?

Linkup repeated relevant results in six requested languages. The French case met the strict rule in 1/2 live rounds. It is a Web retrieval product, not a strict structured financial news feed. Among explicit news APIs in the tested integrations, English was the only language repeated in our fixed company cases.

Which financial news APIs return sentiment?

Finlight exposed a sentiment label alongside structured entities, categories and provenance. Sentiment coverage here measures availability on relevant rows, not sentiment accuracy.

Is there a free financial news API for AI agents?

Several tested providers offer free or trial tiers, but the results above were measured on the plans we tested. Free tiers usually reduce source licensing, history depth and rate limits, so rerun the fixed cases on the exact plan you intend to ship.

How fast are these APIs?

Median end-to-end latency in this run ranged from about 1.6 s (Finlight ticker lookups) to roughly 16 s (Gildata Public Opinion). Treat these as observations from the test window, not as an SLA.

How often should these results be rerun?

Quarterly, and after any material endpoint, source, plan, schema, language, or ticker-routing change. Keep the test cases stable so score changes reflect the product rather than a rewritten benchmark.

简短结论:最佳金融新闻 API 取决于智能体要完成的工作流。在 406 条可发布实测观察中,Finlight 是表现最强的结构化股票代码新闻 API:20/20 次有效调用,并在九个上市市场中的七个市场连续两轮返回相关新闻。Linkup 的多语言检索范围最广,但它属于 Web 检索产品,而不是严格的金融新闻源。EODHD 是覆盖多交易所股票代码的主要备选方案;GildataCaidazi 则更适合中国市场相关工作流。

核心结论速览

AI Agent 的主要需求 建议优先测试 本次实测依据
按股票代码获取结构化新闻 Finlight Financial News 7/9 个市场连续两轮通过;20/20 次有效调用;无效代码控制表现干净
覆盖更多交易所代码 EODHD Financial News 7/9 个市场连续两轮通过;日本和印度市场出现供应商错误
全球公司与宏观新闻发现 Linkup Search 公司场景 18/18、宏观场景 12/12 均返回相关结果
中国内地及香港公司新闻 Gildata Stock News、EODHD 适用场景均为 4/4 条相关结果
中文基金、行业和机构新闻 Caidazi Hybrid Search V2、Gildata Public Opinion 6/6 轮返回相关结果,Top 5 相关度均为 100%
熟悉的英文关键词检索 NewsAPI、Brave News Search 宏观场景分别有 8/10 和 5/6 轮返回相关结果

不存在一个产品可以同时在全球发现、严格股票代码筛选、多语言检索、情绪字段和区域专业新闻中全面胜出。应该根据智能体工作流选择,而不是只看一个总分。

测试范围:7 个金融新闻 API,406 条实测观察

本次综合基准产生了 406 条可发布观察,共执行 410 次实时 API 调用,覆盖 203 个“工具—测试用例”单元,每个单元重复运行两轮。测试场景包含公司、宏观、加密货币、外汇、无效输入,以及中国市场专业新闻,并覆盖九个上市市场和七种请求语言。

股票代码专项测试另包含 120 条已完成观察,覆盖九个上市市场及一个无效代码控制。只有两家供应商达到了多市场排名门槛。某市场只有在两轮测试中都返回目标公司的相关新闻时才算通过;不同供应商使用各自原生代码格式,而不是强行使用同一套代码写法。

所有概览图固定比较七家供应商:Brave、Caidazi、EODHD、Finlight、Gildata、Linkup 和 NewsAPI。名单来自基准版本中至少在一个采购场景达到入选条件的供应商,并为后续复测保持固定。某供应商后续表现下降时仍保留在图中,但会标记为“未达门槛”。

全球公司发现
适用实时轮次中的相关结果重复率
是否达标还取决于公开门槛和样本覆盖。
Linkup
100% (18)
Gildata
89% (18)
EODHD
78% (18)
Caidazi
56% (18)
NewsAPI
38% (16)
Finlight
11% (18)
Brave
100% (4)

Brave 只有四个适用轮次,低于场景资格所需覆盖范围;百分比是带日期的实测观察,不是通用供应商总分。

“API 返回有效响应”“结果形态符合预期”和“新闻内容与目标相关”是三个不同指标。把它们合并成一个成功率,会奖励返回无关兜底内容的接口,也会错误惩罚如实返回空结果的接口。

为了避免全球检索能力掩盖区域或专业场景优势,我们分别比较全球公司发现、中国内地及香港公司新闻、中文专业新闻、宏观及跨资产发现、多市场股票代码五类采购场景。达标 表示供应商的最佳适用工具通过该场景门槛;未达标 表示完成了测试但低于门槛;N/A 表示所测接口契约无法表达该请求。

工作流适配地图
是否适合,取决于智能体要完成的具体任务
达标未达标不适用
供应商全球中国/香港专业新闻宏观Ticker
Brave未达标达标
Caidazi达标达标达标达标
EODHD达标达标达标
Finlight未达标未达标未达标未达标达标
Gildata达标达标达标达标
Linkup达标达标达标
NewsAPI未达标未达标达标
供应商 全球公司发现 中国内地/香港公司新闻 中文专业新闻 宏观/跨资产 多市场股票代码
Brave 未达标;100%(4) N/A N/A 达标;83%(6) N/A
Caidazi 达标;56%(18) 达标;50%(4) 达标;100%(6) 达标;50%(12) N/A
EODHD 达标;78%(18) 达标;100%(4) N/A N/A 达标;7/9 个市场
Finlight 未达标;11%(18) 未达标;0%(4) 未达标;0%(6) 未达标;33%(12) 达标;7/9 个市场
Gildata 达标;89%(18) 达标;100%(4) 达标;100%(6) 达标;75%(12) N/A
Linkup 达标;100%(18) 达标;100%(4) N/A 达标;100%(12) N/A
NewsAPI 未达标;38%(16) 未达标;0%(4) N/A 达标;80%(10) N/A

这是一张资格地图,不是通用质量总分。详细指标会分别展示相关度、新鲜度、字段完整性、来源多样性、重复情况、情绪字段覆盖、延迟和负向控制。综合发现测试使用各产品的宽泛搜索或新闻源输入;严格股票代码测试则使用专用 ticker 参数和供应商各自的代码格式,两类结果不能互换。

评分方法:公开的评测标准

最小评分单位是一条观察,即一个产品、一个适用测试用例和一轮实时运行。“单元”表示一个产品与测试用例的组合,“API 调用”表示一次实际网络请求。我们同时报告三者,避免把评分分母和请求量混在一起。

指标 公开计算方式与通过规则
有效 API 响应率 返回有效成功结构的观察数 ÷ 已完成观察数。有效空结果算 API 成功,供应商错误不算。
预期响应率 正向用例至少返回一篇文章才通过;无效输入控制只有在不返回文章时才通过。
相关结果轮次率 至少有一条 Top 5 结果相关的适用轮次 ÷ 已完成适用轮次。
Top 5 相关度 前五条结果中相关结果的占比;标题、摘要或结构化实体字段出现目标实体或主题别名即视为相关。
新鲜记录率 在用例七天窗口内且具有可用发布时间的记录 ÷ 全部返回记录;缺失时间不获得新鲜度分。
核心字段完整度 标题、URL、发布时间、来源和摘要的实际非空值 ÷ 返回记录中应存在的核心字段值。
可观察来源多样性 每个非空观察中不同规范化来源值的数量;0 表示没有可用来源字段,并不代表供应商没有上游媒体。
重复率 重复的规范化 URL,或标题与日期组合 ÷ 返回记录。
情绪字段覆盖 具有情绪值的相关记录 ÷ 全部相关记录;衡量的是字段可用性,不是情绪判断准确率。
请求语言匹配率 匹配请求语言的记录 ÷ 语言明确或可检测的记录,并同时展示观察分母。
请求国家匹配率 在指定国家用例中,国家元数据匹配的记录 ÷ 具有可观察国家字段的记录。
中位延迟 所有已完成实时观察的端到端耗时中位数;这是测试窗口观察值,不是 SLA。
严格股票代码市场通过 真实市场只有在两轮专用 ticker/symbol 输入中都返回目标公司相关新闻时才通过。
负向控制通过 无效查询或代码不返回文章;无关兜底结果算失败。

N/A 表示所检查接口的契约无法表达请求;未达标 表示适用用例已经完成,但供应商最佳工具没有通过场景门槛。供应商错误、错误语言、过期或无关的兜底数据,以及意外空结果都会保留在底层评估中。每个适用单元必须完成两轮,才能进入排名。

场景排名要求:有效响应率至少 50%,至少 50% 的适用轮次返回相关结果,并且接口契约覆盖至少 50% 的固定用例。多市场 ticker 排名另要求至少 50% 的有效调用,并在九个市场中的至少五个市场连续返回相关结果。这些是资格门槛,不是测试完成后额外加分。

股票代码会在最终测试前按供应商格式进行规范化。语言能力则拆成三类事实:供应商文档声明什么、当前集成暴露什么、固定实时用例实际返回什么。图表与表格来自同一份净化后的计分卡。

这项评测科学、准确吗?

它是一项受控、可复现的产品选型基准,但不是统计意义上具有最终结论的学术研究,也不是长期可靠性认证。固定用例、明确分母、两轮完成门槛、供应商原生 ticker 格式、负向控制和保留的运行证据,使结果可审计并适合辅助 API 选型。

主要限制是样本规模和时间依赖性。两轮测试可以验证测试窗口内的重复性,但无法估算供应商月度 SLA、完整语言库存或长期召回率。新闻库存还会随新闻周期、授权范围、订阅套餐和索引延迟变化。因此我们始终标明测试日期与分母,不给出一个通用总分,并建议每季度以及集成发生重大变化后重新运行。

按智能体工作流选择最佳金融新闻 API

智能体工作流 本次最佳适配 原因
多市场结构化 ticker 新闻 Finlight Financial News 20/20 次有效调用;7/9 个市场连续两轮相关;无效 ticker 行为干净
全球公司发现 Linkup Search 九个固定公司用例共 18/18 条观察返回相关结果
中国内地/香港公司新闻 Gildata Stock News、EODHD、Linkup 各自在四条适用观察中均返回相关结果
中文专业新闻 Caidazi Hybrid Search V2、Gildata Public Opinion 均为 6/6 条相关结果,Top 5 相关度 100%
宏观和跨资产发现 Linkup Search 12/12 条相关结果;NewsAPI 为 8/10
英文新闻发现 Brave News Search 在适用范围内,宏观及跨资产场景 5/6 条相关
多市场 ticker 备选 EODHD Financial News 7/9 个市场连续返回相关结果

中国内地/香港公司新闻的取舍

下表每家供应商只保留一个达标工具。四个工具都覆盖两个固定的中国内地/香港公司用例,并至少在一半实时轮次中返回相关结果。

产品 相关轮次 Top 5 相关度 新鲜记录 核心字段 每个非空轮次来源数 重复率 相关记录情绪覆盖 中位延迟 无效控制
Caidazi Hybrid Search V2 2/4 100% 20% 100% 2.0 0% 0% 1,809 ms 0/2
EODHD Financial News 4/4 100% 0% 60% 0.0 0% 100% 2,194 ms 0/2
Gildata Stock News 4/4 100% 0% 60% 0.0 0% 0% 11,056 ms 2/2
Linkup Search 4/4 55% 0% 40% 0.0 0% 0% 4,042 ms 0/2

EODHD 和 Gildata Stock News 在固定公司用例中相关性很高;Caidazi 暴露了更完整的核心字段和更多可观察来源。来源数为 0.0 表示规范化响应没有可用来源值,不代表供应商没有上游媒体。

中文专业新闻的取舍

这一组要求至少覆盖基金、行业和机构三个固定用例中的两个,并且每家供应商只保留表现最强的达标工具。

产品 观察数 相关轮次 Top 5 相关度 新鲜记录 核心字段 每个非空轮次来源数 重复率 中位延迟
Caidazi Hybrid Search V2 6 6/6 100% 33% 100% 4.0 0% 2,028 ms
Gildata Public Opinion 6 6/6 100% 53% 48% 0.0 0% 24,739 ms

两者六条观察全部相关。Caidazi 更快,规范化字段和来源更完整;Gildata 在新鲜度窗口内的记录占比更高,但延迟明显更高。

宏观和跨资产场景的取舍

产品 观察数 相关轮次 Top 5 相关度 新鲜记录 核心字段 每个非空轮次来源数 重复率 中位延迟 无效控制
Brave News Search 6 5/6 96% 0% 60% 0.0 0% 1,991 ms 0/2
Caidazi Hybrid Search 12 6/12 32% 44% 73% 4.2 8% 2,902 ms 0/2
Gildata Public Opinion 12 9/12 89% 0% 60% 0.0 0% 15,909 ms 2/2
Linkup Search 12 12/12 100% 0% 40% 0.0 0% 6,372 ms 0/2
NewsAPI Everything 10 8/10 70% 35% 100% 2.6 0% 2,826 ms 2/2

新鲜度和字段指标只基于规范化返回记录。无效控制失败表示接口为故意不存在的查询返回了兜底内容;这不会否定正向用例,但生产智能体必须增加相关性校验。

Finlight:适合结构化智能体工作流的 ticker 新闻 API

Finlight 在严格股票代码套件中完成 20/20 次调用,通过两轮无效 ticker 控制,并在美国、香港、日本、德国、法国、印度和西班牙市场的两轮测试中都返回相关新闻。中国内地和巴西市场两轮均为空结果,中位延迟为 1,574 ms。

Ticker 工作流实测
九个上市市场的连续相关结果
两轮均通过空结果或供应商错误
Finlight
西
7 / 9
EODHD
西
7 / 9

“通过”表示两轮固定实时测试均返回目标公司的相关新闻,仅代表本次观察结果,不是永久覆盖承诺。

Finlight 还返回适合智能体使用的结构化字段:标题、URL、发布时间、来源、摘要、语言、国家、情绪标签、分类和展开后的公司实体。其 confidence 字段表示分类置信度,而不是情绪方向,因此未作为情绪数值计分。

金融新闻 API 支持哪些语言?

语言覆盖必须拆成三层:供应商文档声明什么、当前集成允许智能体请求什么、实时测试实际返回了哪些相关语言新闻。

产品 文档或集成暴露的语言控制 实测可重复返回的请求语言
Finlight Financial News 原生 API 每次接受一个 ISO 639-1 语言参数,默认 en;当前集成未暴露该参数 en
Linkup Search 所测契约没有显式语言筛选 deenesjaptzh;法语用例两轮中通过一轮
NewsAPI Everything 暴露 14 个语言代码 en
Brave News Search 当前集成枚举 arbgbncaeneu en
语言工作流实测
可请求语言不等于实测可重复返回
NewsAPI

暴露 14 个语言代码

实测:英文
文档支持选项不等于固定公司用例中能够连续返回多语言相关新闻。
Finlight 集成

结构化 ticker 能力突出

实测:英文
当前集成只暴露英文,原生供应商文档描述了更广的语言控制。

Finlight 文档说明原生 API 索引多语言新闻,并允许每次请求筛选一种语言;因此当前缺失语言控制属于集成层差距,不能据此推断底层 API 只支持英文。NewsAPI 的语言枚举应以其 Everything 端点最新文档为准。

按市场测试股票代码:哪些 API 能稳定解析代码?

产品 连续两轮通过的市场 边界情况
Finlight Financial News 美国、香港、日本、德国、法国、印度、西班牙 无效代码两轮均返回空结果
EODHD Financial News 美国、香港、中国内地、德国、法国、巴西、西班牙 日本、印度和无效代码返回供应商错误

Ticker 支持不等于关键词召回。产品可能找到包含“Toyota”的文章,却无法处理严格的 7203.T 查询;不同供应商也可能为同一上市标的使用不同后缀。接受用户输入代码的智能体,应在调用新闻接口前按供应商规则进行规范化。

国家和请求语言测试结果

每个适用“工具—测试用例”单元调用两次。只有响应结构符合预期、至少返回一条相关结果、所有可观察语言的 Top 5 结果匹配请求语言,并且可用国家元数据匹配目标市场时,该轮才通过。因此 50% 表示两轮中有一轮完整通过,并不等于文章级语言准确率达到 50%。

严格用例规则
市场—语言单元只有全部条件满足才算通过
每个单元运行 2 轮
响应结构符合预期+结果与目标相关+匹配请求语言+可用国家字段匹配=通过

50% 表示两轮完整测试中通过一轮,并不代表文章级语言准确率为 50%。

金融新闻中的国家标签可能表示媒体所在国、公司注册地、上市市场或文章主题。采购方必须先确认供应商字段语义,再把国家过滤当作交易所过滤使用。

如何为 AI Agent 选择金融新闻 API

如果智能体从股票代码开始,并需要结构化实体、情绪标签、分类、来源证据和干净的负向控制,优先测试 Finlight。若需要 Finlight 多语言检索,应先在集成中暴露其原生 language 参数,再重新运行同一套固定用例。

如果更重视覆盖多个交易所代码,并且智能体可以明确处理供应商错误,可选择 EODHD。若多语言召回与引用来源比严格新闻源契约更重要,可选择 Linkup;英文发现可测试 Brave;需要熟悉的关键词 API 和显式语言选择时可测试 NewsAPI,但声明支持某种语言并不保证能返回相关的本地市场新闻。

中国内地和香港公司工作流应比较 Gildata Stock News、EODHD、Linkup 和 Caidazi Hybrid Search V2。Gildata 与 EODHD 在两个固定公司用例中最稳定;Caidazi 的规范化字段和来源多样性更完整;Linkup 则以较弱结构换取更宽泛的检索范围。

中文基金、行业和机构工作流中,Caidazi Hybrid Search V2 更快、结构更完整;Gildata Public Opinion 六轮均返回相关结果且新鲜记录占比更高,但延迟更大。宏观、加密货币和外汇发现中,Linkup 的相关结果重复性最好,其次是 NewsAPI 和 Gildata。Linkup 与 Brave 均未通过无效查询控制,因此必须在接受兜底结果前验证实体或主题相关性。

局限性

这是一组带日期的实时观察,不是对供应商全部新闻库存的永久结论。新闻可用性会随新闻周期、订阅套餐、来源授权、索引延迟和集成结构变化。空结果不能证明某市场在合同上不受支持,只能说明固定查询在测试窗口内没有返回相关新闻。

基准只评价实际调用的端点和参数,不会把同一供应商其他端点的能力自动转移到被测接口上。

供应商如何加入或改进这项基准

供应商可以提交可比较的新闻发现或 ticker 新闻端点,说明市场和语言控制,并提供测试凭据或 QVeris 集成。入选和排名不能购买;契约事实可以根据文档修正,但性能结果只有在固定实时套件重跑后才会变化。

请先检查产品是否已经收录在 QVeris Provider Hub。如需新增收录、修正端点契约或安排复测,请发送邮件至 support@qveris.ai,主题注明“Financial News Benchmark”。为了加快复测,请提供准确端点与稳定结构、支持的 ticker 格式和交易所、语言代码及筛选层级、国家字段语义、来源与授权限制,以及无效输入和空结果行为。

金融新闻 API 常见问题

哪个金融新闻 API 最适合 AI Agent?

不存在单一赢家。本次基准中,Finlight 最适合结构化 ticker 新闻;Linkup 更适合多语言和宏观发现;EODHD 是覆盖多交易所 ticker 的主要备选;Gildata 和 Caidazi 更适合中国市场。应按工作流选择,而不是只看总分。

可以按股票代码搜索金融新闻吗?

可以。多项集成暴露了专用 ticker 或 symbol 输入,但只有 Finlight 和 EODHD 达到多市场排名门槛。市场覆盖和代码写法不同,应把用户输入转换为供应商格式,例如 7203.TTYO:7203

哪个 API 的实测语言覆盖最广?

Linkup 在六种请求语言中连续返回相关结果,法语用例两轮中通过一轮。但它是 Web 检索产品,不是严格的结构化金融新闻源。在本次测试的明确新闻 API 中,固定公司用例只有英文实现了可重复结果。

哪些金融新闻 API 返回情绪字段?

Finlight 在结构化实体、分类和来源证据之外还提供情绪标签。这里的情绪覆盖只衡量相关记录是否提供该字段,不衡量情绪判断是否准确。

是否存在适合 AI Agent 的免费金融新闻 API?

多家供应商提供免费或试用套餐,但上述结果基于本次实际测试的套餐。免费层通常会限制来源授权、历史深度或请求速率,因此应在准备上线的准确套餐上重跑固定用例。

这些 API 有多快?

本次测试的端到端中位延迟,从 Finlight ticker 查询约 1.6 秒,到 Gildata Public Opinion 约 16 秒不等。这些是测试窗口观察值,不是 SLA。

应该多久重新测试一次?

建议每季度重测,并在端点、来源、套餐、结构、语言或 ticker 路由发生重大变化后重新运行。固定测试用例应保持不变,确保分数变化来自产品,而不是改写后的基准。