02 / Collection
Top stories
All stories-
What Changed in SWE-Bench Pro V2?
Scale AI Labs says SWE-Bench Pro V2 refreshes its public task set and adds checks on task validity, agent access and solution grading—while leaving some risks open.
-
GPT-6 Sol and Luna: Pricing, Access, and Reported Gains
OpenAI says GPT-6 Sol and GPT-6 Luna bring lower API prices, improved caching, and reported performance gains over GPT-5.6 models.
-
Claude Opus 5.5: What the Announcement Says
Anthropic presents Claude Opus 5.5 as the first model in a new Claude 5.5 family, with reported Fable...
-
Claude Opus 5.5 Reported in Claude Code 2.1.280
An X post points to a screenshot listing claude-opus-5-5 in an apparent Claude Code 2.1.280 update, but...
-
GPT-6-Sol Ultrafast Mode: What the Post Says
An X post speculates about a possible GPT-6-Sol Ultrafast mode and a speed of 750 tokens per second, but...
03 / Collection
Latest news
All news-
Qwen4 lineup reported at 2026 Yunqi Conference, but release details remain unknown
A report and event photo identify four names in an upcoming Qwen4 family, while leaving the release...
-
OpenAI creates an independent mathematics advisory group
OpenAI says its new mathematics advisory group will advise on emerging results, research standards,...
-
DeepSeek Reportedly Plans 2T and 8T Models on Huawei Chips
A post reports that DeepSeek is training a model described as 2T, planning an 8T model and expecting...
-
Sam Altman Is Scheduled to Brief the UN Security Council on AI
OpenAI CEO Sam Altman is expected to brief an open UN Security Council meeting on AI and international...
-
OpenAI revenue forecast: $101 billion projected for 2029
An X post forecasts at least $80 billion in OpenAI annual recurring revenue, while its attached chart...
-
Claude Code Adds AGENTS.md Support in Version 2.1.277
Claude Code 2.1.277 adds support for checking and using AGENTS.md when a folder has no CLAUDE.md, with a...
04 / Collection
Leaks
All leaks-
GPT-6 Sol, Luna and Astra: What the Azure Report Shows
An unverified X post points to Azure configuration entries for GPT-6 Sol, GPT-6 Luna and GPT-6 Astra...
-
What Is OpenAI’s Rumored Aeon AI Assistant?
A post links the unconfirmed name Aeon to a possible OpenAI assistant, while an excerpt attributed to...
-
What an X Post Says About GPT-6 Sol, Luna and Anthropic 5.5
A single, truncated X-post excerpt reports rumored OpenAI and Anthropic model updates, but does not...
-
Grok 4.7 Reportedly Spotted in OpenCode Zen
A truncated X excerpt reports that Grok 4.7 appeared in the OpenCode Zen harness, but it does not...
-
GPT-6 Sol API Support: What the Report Shows
A report from @LuminaBench points to possible GPT-6 Sol preparation across model and provider...
-
Claude Opus 5.5 Reportedly in Stealth Testing
A post on X reports that Anthropic is testing Claude Opus 5.5 under the reported codename...
05 / Collection
AI releases
All releases-
Tencent Hunyuan Launches Hy Image3.5 Preview
Tencent Hunyuan has announced a Hy Image3.5 preview with text-to-image and image-to-image generation, up...
-
Xiaomi MiMo-V2.6 Pro and Flash: Reported Benchmarks
Xiaomi announces MiMo-V2.6 Pro and Flash with open model weights and reports mixed benchmark results...
-
Grok 4.7: Pricing, Benchmarks, and Availability
xAI presents Grok 4.7 as a model for longer coding and knowledge-work tasks. Here are its reported...
-
Kimi Code Desktop brings multiple coding agents into one workspace
Kimi describes Kimi Code Desktop as a workspace for managing multiple coding agents, running tasks in...
-
Qwen-Image-2.1: Open-Weight Image Generation and Editing
A grounded overview of Qwen-Image-2.1’s documented image-generation, editing, transparency,...
-
Ternary Bonsai 2 27B Reports 98.2% of Qwen3.8’s Score at 5.9 GB
PrismML says Ternary Bonsai 2 27B is a 5.9 GB model based on Qwen3.8 27B, with 98.2% of its reported...






















