We know models cheat. A new benchmark measures how much, and on what tasks.
DeepSeek V4.1-Flash cuts GPU memory for AI agent sessions by 75%, fitting four times as many concurrent sessions on the same ...
DeepSeek 4.1 Flash packs over 500 billion parameters while cutting memory requirements 4x, rivaling Claude Opus 5 Kim K3 in ...
DeepSeek launched its V4.1-Flash AI model with an ultra-low cached input token price of $0.003 per million off-peak.
Foundation-80B-A3B-Base, an 80 billion-parameter system built for both commercial deployment and academic research. The model is now available on Hugging Face under an Apache 2.0 license, meaning ...
Introduction: OpenAI described itself as a 'deployment company'In April 2026, OpenAI's Chief Revenue Officer (CRO), Denise ...
New research accepted for presentation at AMIA 2026 demonstrates the performance of HEALWELL's DARWENâ„¢ AI-powered SMARTSuite, including clinical summarization results that outperformed published ...
1. Context management decides what the model sees at each moment: the right information in, the irrelevant out. As agents ...
HEALWELL AI Inc. had new research accepted for presentation at AMIA 2026 demonstrating the performance of HEALWELL's DARWEN AI-powered SMARTSuite, including clinical summarization results that ...
A new Creative Writing benchmark from Vulsar AI suggests that the most advanced AI models can now outperform amateur human ...
AgentX benchmark results from NVIDIA's AI Infra Summit show the Vera Rubin NVL72 system delivers up to 30 times more AI agent ...
Insilico Medicine has published a Cell cover study introducing LongevityBench, a suite of specialised AI models and an ...