AI Models

New models, benchmarks, capabilities

50 stories

In the last 7 days, AI Models tracked 31 stories — 42% positive, 26% negative, 32% neutral sentiment, averaging 6.7/10 impact.

Stories appear on this page because our classification stage assigned them this category as their primary topic — each story receives exactly one category per niche, chosen from a fixed list, so a story that touches both a funding round and a product launch in the same week sorts into whichever category best matches its dominant subject, not both. This keeps each category page focused on one beat rather than a blend of unrelated developments, and applies the same source-verification standard used across every story on this site. Sentiment measures the directional read of each development for this category specifically, not the tone of the reporting, and impact weights how consequential a development is — regulatory, financial, or operational — rather than how widely it was syndicated across outlets.

Figures are computed live from our source-verified story record — see our methodology for how impact and sentiment are derived.

Bullish 7

LongCat-2.0: China's Trillion-Param Model Trained on 50K Homegrown Chips Rivals Gemini

Meituan's LongCat-2.0 claims to be the first trillion-parameter model trained end-to-end on a 50,000-chip domestic compute cluster, matching Google's Gemini 3.1 pro. This achievement demonstrates that Chinese AI chips can now support large-scale training, intensifying the global AI race and loosening Nvidia's grip on advanced AI hardware.

Verified by 2 sources
Neutral 7

AI's $34.5B Browser Grab: Perplexity Bids for Chrome's 3B+ Users

In a bold AI distribution move, Perplexity's $34.5B bid for Chrome would embed its AI assistant into the world's most used browser. But the deal risks slashing open web investment by 70% and raises antitrust questions about AI companies controlling critical internet infrastructure.

Verified by 2 sources
Neutral 7

Llama AI Will Decide Winners in Meta's New Prediction Market—No Humans Needed

Meta is putting its Llama large language model at the center of Arena, an AI-native prediction market app that generates questions, personalizes bets, and unilaterally settles outcomes. This raises new frontiers for AI trust, bias, and automated governance in a $1 trillion sector.

Verified by 7 sources
Neutral 5

AI Giant’s Institutional Backing Tested: PKO Cuts NVIDIA Stake 13.3% to $56.68M

PKO Investment Management trimmed its NVIDIA position by 13.3% in Q1 2026, while a director also sold shares. These moves raise questions about institutional conviction in the AI leader at current valuations, even as smaller funds add positions. For the AI sector, it signals that even top beneficiaries see profit-taking after a historic rally.

Verified by 2 sources

Source: The Lincolnian Online · Bbns

Bearish 6

80% of OpenAI’s GPT-5.5 Answers Leaned Left in Political Bias Test

A Washington Post investigation exposes significant left-wing political bias in leading AI models, with OpenAI's GPT-5.5 showing an 80% left-leaning response rate. Google's Gemini demonstrates over 90% balanced answers, highlighting feasible neutrality in AI systems. This raises urgent questions for AI developers regarding training data, value alignment, and user trust.

Verified by 10 sources

Source: komonews.com · wcyb.com

Neutral 8

GPT-5.6 Sol & Mythos 5: How Government Vetting Reshapes Frontier AI Release

The Trump administration’s direct intervention in model releases marks a turning point for AI research and deployment. With GPT-5.6 Sol capped at 20 users and Mythos 5 redirected to defensive cybersecurity, the AI community confronts a new era of guarded capability dissemination.

Verified by 31 sources
Neutral 6

How Hong Kong's 2-Language Fluency Exposes AI's Cross-Lingual Hallucination

A firsthand test of Google Gemini revealed a dangerous cross-lingual hallucination: fabricated English citations masking unverified Chinese content, while Chinese outputs dropped global context. Bilingual cross-examination broke the deception, revealing systemic epistemic asymmetry in LLMs—and highlighting why Hong Kong’s dual-language population is a new asset for AI safety.

Verified by 3 sources

Source: Hu Chao (hk) · Hu Chao (cn)

Bullish 9

Amazon Deploying Custom Trainium AI Chips in India as Part of $21B Cloud Bet

Amazon's $21 billion India AI and cloud infrastructure commitment will deploy custom Trainium AI chips and Bedrock managed AI services through expanded AWS data centers in Mumbai and Hyderabad. The investment positions India as a significant node in Amazon's global AI infrastructure network, with startups, enterprises, and government organizations gaining access to cutting-edge AI compute and inference capabilities.

Verified by 6 sources
Neutral 6

Agentic AI's ROI problem: Publishers at Cannes demand proof, not hype

At Cannes Lions 2026, the AI conversation pivots from experimentation to hard ROI, as publishers scrutinize whether agentic media buying and AI-powered search can genuinely improve discoverability and revenue. The tech industry must now demonstrate that agentic systems reduce friction rather than adding another layer of fees.

Verified by 2 sources

Source: digiday.com · Digiday

About AI AI Models coverage

According to our own tracking database, this category has accumulated 320 ai models stories since coverage began. This page aggregates the latest ai models stories within our AI coverage area. Every story is cross-referenced across multiple primary sources, scored for sentiment and operational impact, and timestamped so fresh developments surface first. We track new models, benchmarks, capabilities and surface the angles a domain expert would actually read.

Story selection follows our editorial methodology — impact scoring weights regulatory, financial, and operational developments distinctly. Sentiment is classified across five tiers via supervised classification trained on labeled industry corpora. See our glossary for term definitions and our trends index for longitudinal patterns across the AI beat.

Stories only surface on this page once the classifier scores them at a minimum 35 percent relevance to the category. According to that methodology, reviewed July 2026, this follows multi-source corroboration standards recommended by journalism research bodies such as the Reuters Institute for the Study of Journalism.

See something wrong on this page — a wrong stat, a broken source link, a miscategorized story? Report a data issue.

SignalWhat it tells you
Verified by N sourcesConfidence the story isn't a single-source rumor — N≥2 means the development is independently corroborated.
Impact score (1-10)Estimated regulatory, financial, or operational impact. 8+ indicates a story experienced operators should act on.
SentimentFive-tier classification (very bullish through very bearish) trained on labeled AI-specific corpora.
Time stampRecency. Fresh stories (under 1h) render with a highlighted timestamp; stale stories (≥24h) render dimmed.