►
AI News: Dots, GPT-6.1 Sol, Sonnet 5.5, Gemini 4, and everything you need to know
Matt Wolfe · 2026, ai news, openai, gemini
AI directory search
Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.
45 matches for "2026"
Watch first when you want a fast feel for the topic before opening courses, docs, or profiles.
►
Matt Wolfe · 2026, ai news, openai, gemini
►
OpenAI · 2026, ai agents, openai, dots
►
Sam Witteveen · 2026, openai, agents api, decisions api
►
OpenAI · 2026, openai, devday, gpt-6.1 sol
►
AI Revolution · 2026, manus 2.0, claude sonnet 5.5, ai agents
►
Fahd Mirza · 2026, local models, qwen3.8, hermes agent
►
Silicon Valley Girl · 2026, ai adoption, business workflows, operator education
►
Claude · 2026, claude code, coding agents, frontier models
►
Fireship · 2026, gemini, frontier models, multimodal ai
Connected-finance safety guide · OpenAI · Beginner
OpenAI announced on October 2, 2026 that ChatGPT Finances is rolling out to Free and Go users in the United States. Use this guide to understand the connected-account workflow, check the source data behind answers, and review privacy controls before linking an account.
chatgpt, personal finance, connected accounts, plaid, privacy
Desktop agent safety guide · GitHub · Intermediate
GitHub released Copilot computer use in public preview on October 1, 2026. Use this guide to decide when visual desktop control is appropriate, enable it deliberately, and test its permission and stop controls before real work.
github copilot, computer use, desktop agents, copilot cli, gui automation
Tabular foundation model explainer · NVIDIA · Intermediate to advanced
NVIDIA released Kumo Tabular on September 29, 2026. Use this guide to understand its in-context prediction workflow, reproduce a baseline, and test its vendor-reported results on your own tables.
kumo tabular, tabular foundation models, classification, regression, in-context learning
Model migration explainer · Anthropic · Intermediate to advanced
Anthropic released Claude Sonnet 5.5 on September 28, 2026. Use this guide to understand the efficiency claims, breaking API changes, and a safe migration test from Sonnet 5.
claude sonnet 5.5, claude code, coding agents, model migration, adaptive thinking
Reasoning-efficient coding model explainer · Fireworks AI · Intermediate to advanced
Fireworks released Ember-1 on September 23, 2026 as a Kimi K3-based model trained to use fewer reasoning tokens. Use this explainer to evaluate the quality, cost, latency, and context-growth claims on your own agent workload.
ember-1, kimi k3, reasoning models, reasoning tokens, cost optimization
MCP gateway launch explainer · Google Cloud · Intermediate
Google announced API Gateway MCP support on September 24, 2026. Use this explainer to assess the request flow, security boundary, preview limits, and a safe first rollout.
mcp, model context protocol, api gateway, openapi, ai agents
AI-assisted science explainer · Anthropic · Intermediate
Anthropic reported ART on September 23, 2026. Use this explainer to separate what its agents found, what the lab confirmed, and what remains a hypothesis.
claude, ai agents, ai for science, biology, genome mining
Frontier coding model launch explainer · SpaceXAI · Intermediate to advanced
SpaceXAI released Grok 4.7 on September 21, 2026 for coding, agentic tasks, and knowledge work. Use this guide to compare its task reliability and total cost with your current model.
grok 4.7, coding agents, long-running agents, model selection, evals
Multi-model routing launch explainer · Unbiased · Intermediate to advanced
Union Alpha was revealed as Unbiased's Pareto 26.9 on September 17, 2026: a hosted system that routes work across several models and returns one checked answer.
pareto 26.9, union alpha, model routing, ensembles, coding agents
Omnimodal model launch explainer · Qwen · Intermediate
Qwen3.8-Omni-Flash launched September 18, 2026 for long audio and video analysis, text answers, web search, and tool-calling workflows.
qwen3.8, omnimodal models, audio, video, agents
System One Model launch · TypeSafe AI · Intermediate to advanced
Jev launched on September 15, 2026 as TypeSafe AI's first System One model: a non-chat model that turns application state into typed decisions and probabilities for software workflows.
jev, system one models, rlcd, structured decisions, calibrated probabilities
Coding agent evaluation guide · Google Developers Blog · Intermediate to advanced
You want Google's September 9, 2026 guide to evaluating coding agents with small behavioral checks, outcome-based assertions, and batch runs that catch regressions without treating a single benchmark score as the whole story.
coding agents, evals, behavioral evaluations, regression testing, harness engineering
Managed agent API launch and guide · OpenAI · Intermediate to advanced
You want OpenAI's September 10, 2026 launch guide for building long-running cloud agents with the managed Codex harness, context compaction, tool search, multi-agent delegation, and your choice of hosted or self-managed sandbox.
openai, agents, codex, context compaction, multi-agent
Voice agent model launch · OpenAI · Intermediate
You want OpenAI's September 10, 2026 launch of GPT-Live-1 in the API for full-duplex voice agents, including its $0.05-per-minute front-end pricing and guidance on pairing it with a backend model and agent harness.
openai, voice agents, realtime, api, model selection
Data agent product guide · OpenAI · Intermediate
You want OpenAI's September 10, 2026 overview of the Data agent in ChatGPT Work, which connects company data, investigates changes, and produces interactive dashboards through a conversational workflow.
openai, data agents, analytics, enterprise, agent workflows
AI safety incident assessment · Anthropic · Advanced
You want Anthropic's September 9, 2026 assessment of four incidents in which Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations, including the model behaviors and evaluation-design failures involved.
anthropic, ai safety, cybersecurity, agent evaluation, alignment
Frontier AI safety policy guide · OpenAI · Advanced
You want OpenAI's September 9, 2026 policy proposal on frontier AI standards, independent assessments, incident reporting, and preserving human control as capabilities advance.
openai, ai safety, frontier models, evaluations, governance
Agent security architecture guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 architecture for combining customer-controlled data storage with automated misuse monitoring for sensitive frontier-model workloads.
anthropic, ai safety, enterprise, privacy, agent security
Model launch and safety guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 model announcement covering Fable 5.1 for general availability and Mythos 5.1 for trusted-access research use.
anthropic, claude, frontier models, coding agents, ai safety
Model launch and selection guide · Google AI · Intermediate
You want Google's September 2, 2026 launch details for Gemini 3.8 Flash and Flash Cyber, including their positioning for agentic workflows and cybersecurity.
google, gemini, frontier models, agentic workflows, cybersecurity
Coding agent workflow release notes · Qwen · Intermediate to advanced
You want Qwen Code's September 10, 2026 update on workflow run history, agent and token observability, context-usage inspection, combining Plan with YOLO mode, and named parallel channel tasks in isolated workspaces.
qwen code, coding agents, workflow observability, context management, planning
Agent migration case study · Mistral AI · Intermediate to advanced
You want Mistral's September 9, 2026 field report on migrating 40,000 lines of Fortran 77 to C++, including a parity harness built before migration, agent-generated documentation, planner-coder-tester-reviewer workflows, and the human checkpoints needed for maintainable results.
mistral, coding agents, legacy modernization, parity testing, agent workflows
Agent security architecture guide · Meta AI · Intermediate to advanced
You want Meta's September 8, 2026 technical account of defense-in-depth for a long-running personal agent, including isolated runtime cells, credential surrogates, a separate permission authority, tainted-egress tracking, scoped approvals, browser controls, red teaming, and prompt-injection evals.
meta, muse, agent security, prompt injection, least privilege
Model launch and selection guide · Meta AI · Intermediate
You are comparing frontier models for long-horizon coding or knowledge work and need Meta's September 2, 2026 release notes on Muse Spark 1.3's tool use, instruction following, user steering, multitasking, model self-awareness, and availability through Muse Code and Meta Model API.
meta, muse spark 1.3, frontier models, coding agents, long-horizon agents
Open security benchmark · Hugging Face Community · Intermediate to advanced
You want a reproducible September 5, 2026 benchmark for comparing how agentic models handle indirect prompt injection, with public data, a public harness, control runs, tool-call traces, and outcome metrics tied to unauthorized payment actions.
agents, prompt injection, agent security, evals, tool use
Coding agent evaluation guide · GitHub · Intermediate to advanced
You want GitHub's September 2, 2026 evidence for measuring coding-agent efficiency across the whole task, including selective output compression, preserving useful context, benchmark regressions, and controlled production experiments.
github copilot, coding agents, context engineering, evals, cost optimization
Research report · OpenAI · Intermediate
You want OpenAI's September 6, 2026 evidence and measurement framework for how coding agents are changing research workflows, including task delegation, parallel agent use, capability tracking, and the limits of interpreting productivity signals.
openai, codex, research agents, automated research, agent adoption
Model orchestration research preview · GitHub · Intermediate to advanced
You want GitHub's September 4, 2026 technical explanation of runtime model orchestration for coding tasks, including plan decomposition, draft-critique-revise patterns, model cascading, evaluation design, and quality-versus-cost tradeoffs.
github copilot, model orchestration, model selection, coding agents, evals
Agent infrastructure guide · Anthropic · Intermediate to advanced
You want Anthropic's September 3, 2026 workflow for declaring agents, environments, skills, memory stores, and scheduled deployments in a repository, previewing changes, and applying them reproducibly in CI.
anthropic, ant cli, managed agents, infrastructure as code, skills
Open-source guide and repo · Hugging Face · Intermediate
You want a September 3, 2026 walkthrough of funes, an open-source local memory layer that indexes agent traces, preserves provenance, and lets Claude Code, Codex, pi, and Hermes recall decisions across sessions and machines.
hugging face, coding agents, agent memory, codex, claude code
Coding agent release notes · Qwen · Beginner to advanced
You want Qwen's September 3, 2026 practical guide to its latest coding-agent workflows, including learning-oriented output styles, scheduled tasks, a browser terminal, safer workspace writes, skills, MCP stability, and current vision-model support.
qwen code, coding agents, output styles, scheduled tasks, web shell
Agent product design guide · xAI · Beginner to intermediate
You are designing an always-on agent product and want a concrete September 3, 2026 case study covering persistent identity, memory, status, approvals, artifacts, scheduled routines, event triggers, and agent-to-agent work.
xai, grok bot, persistent agents, agent product design, memory
Agent interoperability standard · Anthropic · Intermediate to advanced
You want Anthropic's August 27, 2026 research preview of a model-agnostic standard for connecting agents to lab and manufacturing hardware through programmable interfaces and protocols such as MCP, with safety evaluation built into the rollout.
anthropic, model hardware standard, agents, mcp, robotics
Model launch · OpenAI · Intermediate to advanced
You want OpenAI's September 3, 2026 launch overview for GPT-6 Astra, including its rollout, computer-use and coding focus, long-context behavior, benchmark disclosures, pricing, and new Codex context-retrieval workflow.
openai, gpt-6 astra, frontier models, coding, computer use
Model docs · Google AI for Developers · Intermediate
You need Google's official September 2, 2026 GA reference for its newest Flash model, including the stable model ID, 1M-token input window, thinking levels, multimodal inputs, tools, and production inference options.
gemini, gemini 3.8 flash, coding, autonomous agents, long context
Model docs · Google AI for Developers · Beginner to intermediate
You want Google's September 3, 2026 preview documentation and examples for generating short clips or full songs from text and image inputs with the Lyria 3.5 Clip and Pro models.
google, lyria 3.5, music generation, audio generation, multimodal