GPTProto

Model Releases

453 picksNewest firstLatest article Oct 2, 2026, 4:00 PM
6 stories
15:44 The Decoder:AI News(RSS)AI score 67/100

Black Forest Labs launches Flux 3 Image with multi-step editing that leaves the rest of your picture alone

Oct 2, 2026 Black Forest Labs has released Flux 3 Image, the image side of its Flux 3 model family. The model supports multi-step edits without changing other parts of the image, BFL claims, and covers text-to-image, image-to-image, text rendering, and photorealism. Users can compose scenes with bounding boxes, include up to ten reference images, and output up to 4K. A free demo is available here. API access is 50 percent off through October 8. Companies can license commercial weights to run and

14:35 MarkTechPost(RSS)AI score 54/100

AWS Strands Labs Releases Strands Decider 2B: An Open Source Decision Model That Picks Options in About 115 ms

AWS Strands Labs releases Strands Decider 2B, an open source decision model. It does not generate text. It reads a state and typed questions, then returns a choice, a yes/no probability, or a score with a calibrated confidence. The model has 1.9 billion parameters and runs locally on a CPU, a consumer GPU, or an Apple silicon Mac. Is it deployable? Yes, for local and self-hosted use. Weights are on Hugging Face under Apache-2.0, and pip install strands-decider gives a CLI and an HTTP server. The

01:33 The Decoder:AI News(RSS)AI score 60/100

Ideogram says its new model can edit part of an image without messing up the rest

Oct 1, 2026 Ideogram says its new model Ideogram 4.5 solves one of AI image editing's biggest headaches. When you edit part of an image, the rest shouldn't change. Swap someone's outfit, and their body shape and background should stay intact. Leading models like GPT-Image 2.5 and Nano Banana have gotten much better at this, but they still tend to produce artifacts, especially after multiple edits. Ideogram 4.5 only touches what the user tells it to, the company claims. Target use cases include p

00:49 TechCrunch:AI(RSS)AI score 58/100

Amazon releases its own Jev clone as decision models flood the web

Amazon Web Services released an open source decision model inspired by TypeSafe’s Jev, with AI developers increasingly seeking intelligence that is more suited to computer automation than frontier LLMs. Amazon’s Strands Decider 2B, released the same week OpenAI announced a similar offering, is a high-speed, low-cost way to sort between pre-decided options and deliver a measure of how confident it is in its choice. The model is fully open sourced, available now, and small enough to run locally. A

5 stories
16:00 Ai2 / Allen Institute for AI(RSS)AI score 58/100

Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs

Today we’re releasing Olmo-core 3, a significant upgrade to our framework for developing large language models featuring a redesigned open mixture-of-experts (MoE) training system.Olmo-core 3 is designed to scale MoE training into the trillion-parameter range while preserving computational efficiency. It’s one of the core systems behind the next generation of Olmo, and part of our ongoing commitment to open up the tools and training infrastructure behind each new model.Training large AI models t

14:57 MarkTechPost(RSS)AI score 62/100

NVIDIA Releases Kumo Tabular: Open Tabular Foundation Models That Predict New Rows in a Single Forward Pass

NVIDIA has released Kumo Tabular, a new family of tabular foundation models (TFMs) for classification and regression. If you have followed TabPFN or TabICL, the setup will look familiar. The model takes labeled rows as context and predicts new rows in one forward pass. There is no training, no hyperparameter tuning, and no feature engineering. Kumo Tabular comes in Small, Medium, and Large versions, spanning about 28M to 215M parameters. It runs through NVIDIA’s open-source structured-data-model

11:23 MarkTechPost(RSS)AI score 53/100

Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence

Perplexity Research and turbopuffer have released pplx-embed-v2-context-9b-preview, a contextual embedding model for RAG pipelines. Each chunk is embedded with the full document in view. The real change is the training signal. The model learns to retrieve the answer along with the context needed to verify it, not one ‘gold passage.’ PIs it deployable? Yes, as a self-hosted preview. Weights are on Hugging Face under the MIT license. Loading requires transformers>=5.4.0 with trust_remote_code=True

04:01 Google DeepMind:Blog(RSS)AI score 78/100

Gemini 4 Argon: our next era of frontier intelligence

Sep 30, 2026 | Gemini 4 Argon delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense. In this article Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program. Built to sustain deep reasoning across complex, long-horizon workflows, Argon is fundamentally changing the way we work and build at Go

5 stories
05:47 MarkTechPost(RSS)AI score 62/100

Liquid AI Releases d1: A Decision Model That Returns Calibrated Probabilities With Zero Output Tokens

Liquid AI has released d1, a decision model built for structured choices instead of text generation. You give it context and a set of typed questions. It returns calibrated probabilities across a fixed set of outcomes in a single call, with zero generated tokens. The target is the work many teams still send to general LLMs: classification, ticket routing, scoring, moderation, reranking and LLM-as-judge checks. Is it deployable? Yes, today, as a hosted API. d1 runs on the Liquid API under the mod

01:20 OpenAI:官网动态(RSS · 排除企业/客户案例)AI score 86/100

Introducing GPT-6.1 Sol

Near-Astra intelligence for a fifth of the price We’re introducing GPT‑6.1 Sol, an upgrade to GPT‑6 Sol that nearly matches GPT‑6 Astra’s intelligence on agentic coding, computer use, and professional work at one-fifth of Astra’s standard input and output token prices. Cached input costs just $0.10 per million tokens—95% less than standard input pricing and 50% less than GPT‑6 Sol’s cached input pricing—giving developers more room to build and run capable agents that reuse context across request

00:00 Artificial Analysis 完整文章(网页)AI score 51/100

Korean AI Lab Upstage releases Solar Mini 4

Korean AI Lab 🇰🇷 Upstage has released Solar Mini 4 which scores 24 on the Artificial Analysis Intelligence Index, but costs ~5x as much per task as GPT-6 Luna (max) despite similar per-token pricesSee model page Upstage has released Solar Mini 4, a new proprietary reasoning model. Upstage reports 35B total and 3B active parameters, setting a new Pareto optimal point on Intelligence Index vs. Active Parameters for models under 3B active parameters. It also scores 16 points higher than Upstage's

00:00 Artificial Analysis 完整文章(网页)AI score 78/100

Gemini 4 Argon: Google is back as one of the top three labs in intelligence achieved

See model page Google’s new Gemini 4 Argon equals GPT-6 Astra on the Artificial Analysis Intelligence Index at 60% of the Cost per Task with discounted prices Gemini 4 Argon is Google DeepMind’s first proprietary model above the Flash class in over 7 months. With high reasoning (the highest available), it scores 53 on the Artificial Analysis Intelligence Index, matching GPT-6 Astra (max, 53) and 1 point ahead of GPT-6.1 Sol (max, 52), with gains driven by lower hallucinations and stronger agenti

5 stories
23:30 Hugging Face:Blog(RSS)AI score 67/100

NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction

Highlights (TL;DR) NVIDIA Kumo Tabular, part of the NVIDIA Kumo Structured model collection, is an open foundation model for tabular data now available on Hugging Face. Given a table of labeled rows, it predicts the labels of new rows in a single forward pass, with no training, no tuning, and no feature engineering, for both classification and regression. It was pretrained only on artificial data, comes in three sizes (28M to 215M parameters), runs through our open-source library, and is release

22:45 The Decoder:AI News(RSS)AI score 65/100

ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

Elevenlabs is releasing Eleven v4, a new speech model that follows direction cues more accurately and keeps voices consistent across long productions. A new model architecture also powers the Turbo variant for real-time voice agents. Eleven v4 generates laughter, whispers, and sounds like slamming doors more reliably than its predecessor. The Turbo variant for voice agents starts producing speech in about 150 milliseconds. Eleven v3, released just over a year ago, already supported these audio t

18:00 OpenAI:官网动态(RSS · 排除企业/客户案例)AI score 85/100

DevDay 2026 Recap

DevDay 2026 is our biggest yet, with more than 20 major announcements across ChatGPT, Codex, our models, and entirely new forms of working with AI. We believe AI can help bring about a new renaissance of creativity and discovery. It should give people more time for what matters to them, more freedom to pursue their ideas, and the ability to do things they didn’t think were possible. Today, we introduced agents that can take on ongoing responsibilities and new ways for people and AI to work toget

15:38 MarkTechPost(RSS)AI score 67/100

H Company Releases Holo4: Open-Weight Computer-Use Models That Click, Code and Call Tools Across Desktop, Web, Android and APIs

H Company has released Holo4, a family of generalist computer-use models for AI agents. One set of weights clicks and types on screens. It also writes code and calls MCP or API tools. Holo4 ships in 2 sizes: Holo4 27B (dense) and Holo4 35B-A3B (Mixture of Experts, 3B active). Both serve a 256K context on the H Models API. Is it deployable? Yes. Holo4 35B-A3B ships Apache 2.0 weights for commercial self-hosting. Holo4 27B weights are CC BY-NC 4.0, so commercial use of 27B runs through the H Model

01:58 Anthropic:Newsroom(网页)AI score 88/100

Introducing Claude Sonnet 5.5

Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Claude Sonnet 5, runs 30%+ faster, and costs up to 30% less for most work.Sonnet 5.5 is a faster, lower-cost complement to Claude Opus 5.5. Where Opus 5.5 is built for complex work requiring careful judgment, Sonnet 5.5 is strongest at well-scoped everyday tasks, fixing bugs, and creating polished documents, slides, and spreadsheets. It’s also got a sharp eye for design. Claude Haiku 5.5, built fo

4 stories
20:00 ElevenLabs:Blog(网页)AI score 75/100

Introducing Eleven v4, our most emotive model

A line of text can change significantly depending on how it’s spoken. “I need you to stay calm" should sound different depending on who's saying it, whether that's a doctor delivering it gently to a frightened patient, or a character in a game shouting to his squad before dropping into battle.Today we're launching Eleven v4, our most emotive text-to-speech model yet, and its low-latency variant, Eleven v4 Turbo.Ranked #1 by Artificial Analysis1, and preferred by ~75% of listeners in blind head-t

20:00 Anthropic:Claude.dev 开发者博客(RSS)AI score 87/100

Building with Claude Sonnet 5.5

Claude Sonnet 5.5 is our second model in the Claude 5.5 family after Opus 5.5. It's a clear upgrade over Sonnet 5 and is smarter, more efficient and 30% faster. The per-token price is unchanged and because Sonnet 5.5 typically needs far fewer tokens to do the same work, it costs up to 30% less for most work. FIG ALance Martin's code-to-painting demo: each model writes code that repaints the same photograph. From left: the photograph, Claude Sonnet 5, Claude Sonnet 5.5 and Claude Opus 5.5. Credit

17:44 Hugging Face:Blog(RSS)AI score 71/100

Holo4: powering generalist computer-use agents

Holo4 is our new series of agentic models. It comes in two sizes: 27B dense and 35B-A3B Mixture of Experts. Both are available on the H Models API. We are also releasing an updated version of Holotron 3: Holotron4 Nano. Holo4 builds on our previous model and interacts with software through any available interface: GUIs, code, MCP and APIs. It scores well on academic benchmarks, but we built it for real business workflows. It was trained through supervised and reinforcement learning on a large se

08:00 Claude Platform:开发者版本说明(RSS)AI score 71/100

Claude Platform release notes — September 28, 2026

The Claude Platform release notes list changes to the Claude API, the client SDKs, and the Claude Console, newest first. September 30, 2026 We announced the deprecation of the Claude Sonnet 4.5 model (claude-sonnet-4-5-20250929), with retirement on the Claude API scheduled for November 30, 2026. We recommend migrating to Claude Sonnet 5.5. Read more in Model deprecations. September 28, 2026 We've launched Claude Sonnet 5.5 (claude-sonnet-5-5). It's available on the Claude API, Claude in Amazon B

3 stories
1 story
1 story
1 story
2 stories
1 story
1 story
1 story
1 story
4 stories
1 story
1 story
1 story
2 stories
3 stories
5 stories
1 story
3 stories
2 stories
3 stories
1 story
1 story
2 stories
1 story
1 story
3 stories
2 stories
5 stories
10 stories
21:51 LMSYS:Blog(Chatbot Arena 团队)AI score 74/100

Blog SGLang Adds Day-0 Support for NVIDIA Nemotron 3.5 Lightning SGLang is excited to announce Day-0 support for NVIDIA Nemotron 3.5 Lightning, a customizable open model built to power always-on agents across local systems, the edge, the datacenter, and the cloud. ... NVIDIA Nemotron Team and SGLang Team

SGLang 宣布对 NVIDIA Nemotron 3.5 Lightning 提供 Day-0 支持,该开源模型为 30B 总参数、3B 激活参数的混合专家架构,支持最长 1M token 上下文,可从 Hugging Face 下载 BF16 和 NVFP4 权重。模型支持 MTP、DFlash、DSpark 三种投机解码技术,并可通过 OpenAI 兼容 API 接入智能体工作流。

20:10 蚂蚁 inclusionAI:HuggingFace 新模型AI score 64/100

inclusionAI/Ling-3.0-flash-base-midtrain

蚂蚁 inclusionAI 开源 Ling-3.0 系列语言基座模型,采用高度稀疏(1/64)MoE 架构,512 个路由专家中每 token 仅激活 8 个,总参数量 124B,激活参数仅 5.1B。该系列原生融合线性注意力,并以加权检查点合并替代传统学习率衰减。此次发布包含预训练、中期训练及合并(WSM)等多个训练阶段的检查点,支持继续预训练与微调,模型采用 MIT 许可证。

4 stories
19:51 LMSYS:Blog(Chatbot Arena 团队)AI score 72/100

Blog SGLang Adds Day-0 Support for Muse Glimmer, a Multimodal Model Built for Local Agentic Workflows We're excited to partner with Meta Superintelligence Labs to bring Day-0 support for Muse Glimmer to SGLang, with dedicated optimizations tailored for high-performance inference of agentic workflows o... Meta Superintelligence Labs and the SGLang Team

SGLang 与 Meta Superintelligence Labs 合作,为 30B 参数多模态模型 Muse Glimmer 提供 Day-0 支持,该模型拥有 128k+ token 上下文窗口。

07:58 MarkTechPost(RSS)AI score 75/100

NVIDIA Releases NemotronLabs VoiceChat 11B: An Open Full-Duplex Speech-to-Speech Model with ~450 ms Turn-Taking and Live Tool Calling

NVIDIA 发布开源端到端全双工语音对话模型 NemotronLabs VoiceChat 11B,在统一网络中完成流式语音理解与生成,实测轮换延迟 448 毫秒。该模型为首个支持对话中工具调用的开源全双工模型,通过独立输出通道及预置“保持”话术避免 API 执行期间冷场。权重与容器已公开,但仅限研究用途,需单张 80 GB 显存 GPU,目前无托管 API。

1 story
2 stories
2 stories
2 stories
Showing the latest 100 of 453 stories.