GPTProto

AI Coding News

Latest reporting, research and product updates filed under AI Coding.

450 picksNewest firstLatest article Oct 2, 2026, 1:01 AM

Latest stories

2 stories
4 stories
Anthropic:Claude.dev 开发者博客(RSS)AI score 68/100

Getting started with Claude Code mods

Claude Code already lets you change a lot about how it behaves: settings, permission rules, slash commands, skills and a status line. Mods go further. Mods can rewrite or replace what Claude Code does, and can even draw custom UI. Under the hood, mods are hooks, and they ship inside plugins. Each one is a small JavaScript or TypeScript module that runs inside your session and sees every event as it happens. That makes mods a way to fit Claude Code to how you work. You can add a readout you check

OpenRouter:Announcements(RSS)AI score 67/100

How to Gate Pull Requests on LLM Evals in CI

Changing one line in a support agent’s system prompt can ship an agent that tells customers the refund window is 30 days when your policy says 14. Nothing in a normal CI pipeline checks what the model says, so the build passes and the first person to see the wrong answer is a customer.Gating a pull request on a fixed eval set works the same way as gating on a failing unit test. You keep test cases in the repository, run them when a prompt changes, and block the merge when too many fail.In this g

Google DeepMind:Blog(RSS)AI score 78/100

Gemini 4 Argon: our next era of frontier intelligence

Sep 30, 2026 | Gemini 4 Argon delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense. In this article Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program. Built to sustain deep reasoning across complex, long-horizon workflows, Argon is fundamentally changing the way we work and build at Go

Claude:Blog(网页)AI score 76/100

Customize Claude Code with mods

Today we're introducing mods, small TypeScript functions that change how Claude Code works. A mod can rewrite a prompt, add new UI, replace a built-in feature, or add entirely new functionality. You can write a mod yourself, or ask Claude Code to write one for you. Mods ship inside plugins, so you install and share them like any plugin. They work in the Claude Code CLI and desktop app.Mods run with the same access to your machine as Claude Code itself. They aren’t sandboxed, and you should only

2 stories
6 stories
Pragmatic Engineer(RSS)AI score 78/100

Why has Shopify dropped React Native?

Before we start: given this article is about native mobile development, I want to offer my 2021 ebook, ‘Building Mobile Apps at Scale: 39 engineering challenges’, for free to all readers. (normally costs $20). The book remains relevant on the challenges to solve for large-scale mobile applications, and lists the technologies covered below in this article, Kotlin Multiplatform included.Claim your free copy hereThis offer is valid until Friday, 2 October. On checkout, simply select “The ebook: PDF

OpenAI:官网动态(RSS · 排除企业/客户案例)AI score 85/100

DevDay 2026 Recap

DevDay 2026 is our biggest yet, with more than 20 major announcements across ChatGPT, Codex, our models, and entirely new forms of working with AI. We believe AI can help bring about a new renaissance of creativity and discovery. It should give people more time for what matters to them, more freedom to pursue their ideas, and the ability to do things they didn’t think were possible. Today, we introduced agents that can take on ongoing responsibilities and new ways for people and AI to work toget

Every:最新文章(网页)AI score 80/100

Vibe Check: OpenAI DevDay 2026

OpenAI’s DevDay just kicked off, and we’ve got the rundown of the 20-plus(!) products and features the company announced. Here’s what matters, what it means, and what I learned from my hands-on testing.The quick takeOpenAI wants ChatGPT to become your operating system for work—documents, slides, and agents, all in one place on your computer. It also wants to let developers and startups build businesses on top of it.Five announcements from today show what that looks like: Dots: Persistent agents

Databricks:Blog(RSS)AI score 65/100

How Databricks rolls out frontier models to 12,000 employees on Day 1

Providing our employees access to frontier AI capabilities is a top priority at Databricks, and consequently, it is important to us for them to use new models instantly when they become available. At the same time, it is nontrivial to give more than 12,000 people rapid access to a new model because:Models that are marketed as frontier often aren’t. For example, Opus 5.0 was more expensive and ranked lower on both quantitative and qualitative quality scores among our engineers compared with Opus

Anthropic:Newsroom(网页)AI score 88/100

Introducing Claude Sonnet 5.5

Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Claude Sonnet 5, runs 30%+ faster, and costs up to 30% less for most work.Sonnet 5.5 is a faster, lower-cost complement to Claude Opus 5.5. Where Opus 5.5 is built for complex work requiring careful judgment, Sonnet 5.5 is strongest at well-scoped everyday tasks, fixing bugs, and creating polished documents, slides, and spreadsheets. It’s also got a sharp eye for design. Claude Haiku 5.5, built fo

Artificial Analysis 完整文章(网页)AI score 69/100

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence

See model page GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter of the Cost per Task. Pricing matches GPT-6 Sol at $2/$10 per million input/output tokens, except that the cache read discount rises from 90% to 95%. GPT-6.1 Sol’s overall blended price for agentic workloads is therefore slightly lower than GPT-6 Sol. This represents an additional price cut, following GPT-6 Sol’s original 50% discount from GPT-5.

5 stories
Anthropic:Claude.dev 开发者博客(RSS)AI score 67/100

Automating eval design and hillclimbing with Claude

Evaluations provide a signal on how your app or skill is performing on specific tasks. But designing evaluations, and improving performance on them without fooling yourself, is hard. We've added guidance for both to the claude-api skill. With the skill, you can run /claude-api build-eval to build an evaluation inside your codebase, and run /claude-api hillclimb to improve your application against it, one change at a time, with a held-out set of examples to catch overfitting. In this article, we

Anthropic:Claude.dev 开发者博客(RSS)AI score 87/100

Building with Claude Sonnet 5.5

Claude Sonnet 5.5 is our second model in the Claude 5.5 family after Opus 5.5. It's a clear upgrade over Sonnet 5 and is smarter, more efficient and 30% faster. The per-token price is unchanged and because Sonnet 5.5 typically needs far fewer tokens to do the same work, it costs up to 30% less for most work. FIG ALance Martin's code-to-painting demo: each model writes code that repaints the same photograph. From left: the photograph, Claude Sonnet 5, Claude Sonnet 5.5 and Claude Opus 5.5. Credit

Hacker News:AI 热帖AI score 86/100

Prompting Claude Opus 5.5

This guide covers the prompting patterns specific to Claude Opus 5.5. For the model's capabilities and API changes, see What's new in Claude Opus 5.5. For techniques that apply across all current Claude models, see Prompting best practices. Claude Opus 5.5 generates output tokens more than 30 percent faster than Claude Opus 5 and tends to finish the same task with fewer tokens. Existing Claude Opus 5 prompts should perform well without changes, and the patterns in Prompting Claude Opus 5 remain

Claude Platform:开发者版本说明(RSS)AI score 71/100

Claude Platform release notes — September 28, 2026

The Claude Platform release notes list changes to the Claude API, the client SDKs, and the Claude Console, newest first. September 30, 2026 We announced the deprecation of the Claude Sonnet 4.5 model (claude-sonnet-4-5-20250929), with retirement on the Claude API scheduled for November 30, 2026. We recommend migrating to Claude Sonnet 5.5. Read more in Model deprecations. September 28, 2026 We've launched Claude Sonnet 5.5 (claude-sonnet-5-5). It's available on the Claude API, Claude in Amazon B

Fireworks AI(网页)AI score 62/100

Introducing FireRouter with Opus

Fireworks Nexus enables engineering teams to drop leading open models in the harnesses they already use and cut spend in half without sacrificing speed or quality. The solution includes FireRouter, the first cache-aware router on the market, which makes a big difference in speed and cost. Today, we’re introducing FireRouter with Opus, optimized for the Opus family and now available in both our CLI and, for the first time, as a standalone router model. Any Fireworks account can point to it as a s

2 stories
4 stories
Cognition 模型 / Devin 博客(网页)AI score 72/100

Cognition Crosses $1B in Annualized Revenue Run Rate

Today, Cognition crossed $1B in annualized revenue run rate.We started Cognition in January 2024 because the world needs far more software than it can build. Every company is now a software company, from banks to automakers to governments, and every one of them has more to build than time to build it.Less than two years after Devin became generally available, Devin works alongside engineering teams at GE Aerospace, Rivian, Rohlik, Exa, and many more. We asked a few of them to share what that loo

3 stories
3 stories
3 stories
1 story
2 stories
1 story
2 stories
2 stories
1 story
1 story
1 story
3 stories
1 story
1 story
1 story
2 stories
1 story
Showing the latest 54 of 450 stories.