OpenAI disrupted a Cambodia-based scam operation using ChatGPT to support investment, romance, gambling, and impersonation schemes.
Firehose
Filtered to tagged “open-weight models” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives
Browse by tag
First, abstracting from all the words in my long analysis on OpenAI Astra yesterday, the overall point was that Astra may not be the breakthrough that OpenAI wants you to think it was.
Interconnects has launched an Artifacts Hub and Adoption Dashboard to provide free, curated data on the open model ecosystem, including inference tokens, model intelligence, and adoption metrics. The Adoption Dashboard offers insights into download and derivative model numbers by geography and organization, highlighting the US-China gap and growing players. The data is available for free to support the growth of the open ecosystem. AI summary
Several open models have been released, showcasing their utility on the Pareto frontier, including Inkling by Thinking Machines, Hy3 by Tencent, Laguna S2.1 by Poolside, and Kimi K3 by Moonshot AI, which demonstrate improvements in performance and efficiency. These models are pushing the boundaries of what is possible with open models, with some companies like Thinking Machines and Tencent generating hundreds of millions in revenue per year from their open model finetuning services. The increasing adoption of open models is expected to continue, with potential implications for the AI industry. AI summary
An internal version of OpenAI's Astra model family has solved 10 major open problems in mathematics, quantum complexity, and theoretical computer science, demonstrating significant advancements in scientific reasoning. This breakthrough is attributed to Astra's capabilities, which have already enabled substantial progress in math and science. The solved problems include long-standing open issues in mathematics and theoretical computer science. AI summary
The June edition of my sponsors-only monthly newsletter is out. If you are a sponsor (or if you start a sponsorship now) you can access it here . This month: Accidental cyberattacks by OpenAl and Anthropic models under test GPT-5.6 Sol, Ter…
OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
A full-stack approach to making advanced AI more capable, more affordable, and more widely useful.
DeepSeek V4 Flash now runs on updated weights by default on AI Gateway, with notably stronger agentic capabilities. On Terminal-Bench, it scores 82.7, up 25.8 points from 56.9 in the April preview. Requests to deepseek/deepseek-v4-flash pic…
Explore lower GPT‑5.6 pricing for Luna and Terra—and how OpenAI’s more efficient models help enterprises deploy AI workflows at scale.
**OpenAI** aggressively cut prices for **GPT-5.6 Luna** by 80% and **Terra** by 20%, introducing a faster **Sol Fast** tier with up to 2.5× lower latency at double the price, improving agent workflow costs by roughly 10×. The **ARC-AGI-3** …
Inkling Small from Thinking Machines is now available on AI Gateway. Inkling Small reaches performance comparable to the larger Inkling model at about a quarter of the size, using much less compute per task. It is a broad generalist with na…
avatarin uses OpenAI’s GPT-Realtime to give Yamada Denki shoppers 24/7 multilingual support. In two weeks, 30,000 people used the agent and 92% of survey responses were positive.
Anthropic is opposed to a blanket ban on open weight models, but wants to ban conditions that make them competitive, such as distillation, which it claims allows Chinese companies to evade chip bans and build better models than the US. This approach is seen as protectionist industrial policy that would harm the open AI ecosystem and favor American companies like Anthropic, which has not released any open weight models. AI summary
How two API settings improved GPT-5.6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction.
GPT-5.6 improves AI efficiency across models, inference, and agentic workflows, helping deliver more useful intelligence per dollar.
Subagents get justified by time saved and parallel execution, but Rahul Garg explains that's not what matters most. Every token in the orchestrator's context is competing for its attention, and the real value of a subagent is what it keeps …
You argue that open-weight model releases are a relatively mild AI risk compared with dangers originating inside labs, contending that Amodei’s post mislocates the primary threat. Citing incidents like the OpenAI/Hugging Face episode, you emphasize that modalities of internal development and deployed systems—not merely public model availability—are where the first serious AI incidents are likely to arise. AI summary
Kimi K3, an open-weights model, has been released by Moonshot AI, with a 2.8T-parameter MoE model and 104B active parameters, achieving a 2.5x scaling-efficiency improvement over K2. The model's early evaluations are strong, particularly in agent/coding tasks, with top rankings on Agent Arena and Frontend Code Arena. AI summary
**Moonshot** released the **Kimi K3**, a **2.8T-parameter MoE** model with **104B active parameters/token**, featuring innovations like **Kimi Delta Attention (KDA)**, **Gated MLA**, and **LatentMoE**. The release includes infrastructure co…