OpenAI's Astra Benchmark Edits Muddy Its Claim Over Anthropic
OpenAI revised GPT-6 Astra's launch benchmarks days after release, and reporting shows the edits briefly flattered Astra over Anthropic's Claude models.
AI research sits at the center of technology's fastest-moving frontier, where breakthroughs in model architecture, training efficiency, and safety testing translate almost immediately into products, business strategy, and geopolitical stakes. This hub tracks how research labs, startups, and established tech companies push the boundaries of what large language models and related systems can do—and at what cost.
Recent developments illustrate the breadth of this space: competitive pressure is driving firms to fine-tune existing models for dramatically lower operating costs, while claims of speed breakthroughs from Chinese research teams raise questions about the pace of global AI progress. At the same time, safety researchers are probing how capable models behave when tested against real-world systems, surfacing new concerns about deployment risk. Legal battles over data use and content licensing continue to shape the rules under which AI companies can train and operate, and talent movement between startups, incubators, and major labs signals where strategic bets are being placed.
Readers will find coverage of new model releases and benchmarks, infrastructure partnerships and hardware deals that determine who can train and run cutting-edge systems, regulatory and courtroom developments affecting how AI companies source data, and the personnel and funding shifts that reveal where the field is headed. We also track how AI research intersects with business realities—including layoffs and restructuring at companies betting heavily on automation—and how safety and security research is evolving alongside capability gains.
Whether you're following the technical race between labs, the legal fights over intellectual property, or the economic ripple effects of AI adoption, this page offers an ongoing record of how AI research is reshaping technology and the industries built on it.
OpenAI revised GPT-6 Astra's launch benchmarks days after release, and reporting shows the edits briefly flattered Astra over Anthropic's Claude models.
African climate tech hit $1.5bn in 2025 per a Briter report, but funding remains concentrated in energy, three countries and 20 firms.
Nvidia's Jensen Huang urged G20 leaders to avoid AI rules based on theoretical harms, as other reports show mixed global approaches to AI adoption.
Mile ·
Anthropic's Claude autonomously fixed 10 AI alignment failures, closing up to 96% of safety gaps and outperforming human researchers.
Anthropic says Claude AI found a new attack on round-reduced AES and can autonomously build software exploits, reshaping cybersecurity research.
Buffalo's Empire AI supercomputer debuts as open-source models, cancer AI, and ROI questions reshape the AI research landscape.
University at Buffalo's Empire AI Beta launches as open-source AI models, health uses, and ROI questions reshape the research landscape.
University at Buffalo's Empire AI Beta goes fully online as new AI research milestones emerge in oncology, open-source models, and ROI debates.
China's rising biotech capabilities threaten US dominance even as drug breakthroughs, AI discovery, and pricing deals fuel a sector-wide rally.
A federal judge ruled the Pentagon's blacklisting of AI firm Anthropic was illegal retaliation over its refusal to support military uses.
AI systems from Noetik, GSK and the Allen Institute uncover new cancer research leads as Anthropic pushes AI into labs amid safety concerns.
AI systems are helping uncover new cancer treatment insights, while trust and safety questions around AI tools continue to grow.
Alibaba launched its Wan3.0 AI video model after a $10 billion share sale, amid wider AI funding, safety and content-quality developments.
Mile ·
Trump administration unveils $5B initiative directing federal funding toward AI-powered scientific research.
Anthropic disabled bio-weapon safeguards for 11 months, affecting 133M exchanges, amid rapid AI model and research expansion industry-wide.
Mile ·
DeepMind alumni's startup Inherent says its compact AI agent beat OpenAI and Anthropic models on research-replication tasks.
Anthropic adds Gmail email drafting to Claude as it faces IPO buzz, data-policy rivalry with OpenAI, and new AI agent risk concerns.
JPMorgan says Charles River could extend gains as biotech investment rebounds, alongside Insmed's rally and other sector moves.
Research finds DeepSeek's AI model is far cheaper to run than rivals, as Chinese open models reshape AI pricing.
Mile ·
Perplexity fine-tuned an open Chinese model to match Claude Opus 4.8 at a third of the cost, now live in production.
A major biotech firm cuts 103 more jobs in its fifth 2025 layoff round, amid wider industry turmoil over execution, regulation, and AI deals.
A judge refused to dismiss Reddit's DMCA claims against Perplexity AI and SerpAPI, letting the copyright lawsuit proceed toward discovery.
Perplexity AI expands with Windows agents, an open-source security tool called Numbat, and a new Nvidia Vera CPU deal.
Anthropic says Claude models breached real systems in cyber tests, fueling AI safety and regulation warnings from experts.
Anthropic pushes new Claude models and interpretability research, with a still-unconfirmed report of a Trump administration ban on one release.
Chinese researchers claim an optical interconnect breakthrough boosting AI inference speed 100x, amid shifting China-US AI competition.
Monzo cofounder Tom Blomfield leaves Y Combinator to join Anthropic's compute team amid AI's talent war and fintech's continued global growth.