AI News · Week 17, 2026
Claude Opus 4.7 takes on longer coding tasks as Codex gains Mac controls
· 10 stories · 34 sources
Written with AI, sources linked for every story
Claude Opus 4.7 is designed for longer coding tasks and more reliable checks of its own work. Codex can now click and type in the background on a Mac. Perplexity’s Personal Computer also carries out tasks across files and apps. Alibaba has released Qwen3.6-35B-A3B with open weights, while Anthropic is testing Claude Design for visual work.
Anthropic releases Claude Opus 4.7 for complex coding tasks
Claude Opus 4.7 is designed to handle longer coding tasks and check its own work more reliably. It is available through Claude, the API, and cloud platforms.
Anthropic released Claude Opus 4.7 on April 16, 2026. The model is generally available in Claude products, through the Anthropic API, and from several cloud providers.1 Anthropic says it improves coding, long-running agent tasks, instruction following, and verification of its own work.1 The company calls it its most capable Opus model to date.1
For companies, performance is only part of the decision: a new model also needs to fit existing development workflows. The reported improvements come from Anthropic; teams should test how reliably Opus 4.7 works in their own codebases.1 The model also includes safeguards against prohibited or particularly high-risk cybersecurity requests.2 Organizations using AI for security work should check how those limits affect their workflows before deploying it widely.
What it means for companies
If your company uses AI for software development, test Opus 4.7 on longer tasks from your own codebase. Compare output quality, cost, and any limits on security work against your current model.
OpenAI adds desktop computer use and app integrations to Codex
Codex can now click and type in the background on a Mac. OpenAI is also expanding its integrations and pushing for enterprise adoption.
OpenAI expanded Codex on April 16 for desktop app users signed in with ChatGPT.1 The coding agent can now operate a computer in the background, clicking and typing while the user continues working.1 An in-app browser and 90+ plugins and integrations connect it to more applications and tools.1 OpenAI also points to testing and front-end iteration as tasks beyond writing code.1
On April 21, OpenAI introduced Codex Labs and partnerships with systems integrators to support enterprise adoption.2 Partners include Accenture, Capgemini, and Infosys.2 The update broadens Codex’s role from producing code to carrying out tasks across applications.1 Availability remains a consideration for companies: not every feature is launching at the same time in every region or for every account type.3
What it means for companies
If your development team uses Codex, test computer actions in a limited environment and review integration permissions. For EU teams, plan around delayed access to browser integration and personalization.
Alibaba releases Qwen3.6-35B-A3B with open weights
The new Qwen model activates only part of its parameters for each token. Its weights are available to download, alongside chat and API access.
Alibaba introduced Qwen3.6-35B-A3B on April 15, 2026, as the first Qwen3.6 model with open weights.1 It uses a sparse mixture-of-experts architecture that activates only a portion of its parameters for each token.1 Alibaba also offers the model through Qwen Studio and an API.1 The company positions it for agentic coding and multimodal reasoning.1
The architecture is intended to reduce compute per token compared with a similarly sized model that activates all its parameters.1 That does not establish a fixed savings figure for deployment: memory requirements and infrastructure also matter. Companies can run the weights themselves or try the model through the API.12 The less expensive option will depend on their workloads and infrastructure.
What it means for companies
If you self-host models, assess memory needs and infrastructure costs alongside compute. Compare the self-hosted model with the API using your own workloads.
Anthropic introduces Claude Design for slides and prototypes
Claude Design creates visuals through conversation and exports them to formats including PPTX and PDF. Anthropic is offering it as an experimental preview.
Anthropic introduced Claude Design on April 17, 2026, as an experimental preview from Anthropic Labs.12 The tool creates visuals through conversation, including presentation slides, prototypes, and marketing materials.23 Users can refine results with comments or direct edits, then export them to Canva, as PPTX or PDF, or to Claude Code.23 Access is rolling out gradually to paid subscribers.12
Claude Design is an application powered by Claude Opus 4.7, not a separate new foundation model.12 It extends Anthropic’s offering into visual creation and puts the company in competition with established design tools.24 For businesses, a practical question is whether exported drafts fit existing editing and review workflows. Because the product remains an experimental preview, teams should check layouts and content before publishing them.12
What it means for companies
If your team creates slides or marketing materials, test Claude Design for initial drafts. Check the content, layout, and editability of exported files before sharing them.
Sources (5)
- 1 Anthropic launches Claude Design, a new product for creating quick visuals | TechCrunch techcrunch.com
- 2 Anthropic launches Claude Design tool | VentureBeat venturebeat.com
- 3 Introducing Claude Design by Anthropic Labs x.com
- 4 Anthropic debuts Claude Design, because who needs designers? theregister.com
- 5 Anthropic launches Claude Design following Opus 4.7 model upgrade - 9to5Mac 9to5mac.com
Cloudflare introduces Agent Memory in private beta
The managed service is designed to make information from AI agent conversations available for later tasks. Cloudflare has announced it as a private beta, not a general release.
Cloudflare introduced Agent Memory on April 17 as a private beta. The managed service is designed to extract information from AI agent conversations and retrieve it for later requests without loading the full conversation history into a model’s context window. Agents built on Cloudflare Workers can access it through a binding; agents elsewhere can use a REST API. Cloudflare did not announce general availability.1
Cloudflare positions Agent Memory as part of its infrastructure for AI agents. The service retrieves stored information when needed rather than sending an ever-growing conversation history to a model. That approach could simplify context management for long-running agents, but the announcement does not establish how well it performs in practice. Cloudflare has not provided performance figures, pricing, or usage limits.1
What it means for companies
If you run agents that handle recurring tasks, test whether stored memories can replace lengthy conversation histories. Treat this as an evaluation for now, and check data handling, costs, and usage limits before planning production use.
Perplexity brings Personal Computer to Mac
Personal Computer can carry out tasks across files, browsers, and apps through Perplexity’s Mac app. Access started with Max subscribers.
Perplexity introduced Personal Computer for its Mac app on April 16, 2026. It extends Perplexity Computer and can carry out tasks across local files, browsers, and native apps.1 It can search, read, and edit files and work with apps including Messages, Mail, and Calendar.2 Access began for Perplexity Max subscribers; the company’s changelog says all Max subscribers have access.13
The release adds actions on a user’s Mac to Perplexity’s earlier cloud-based system.1 That makes it relevant to tasks requiring information from several apps, rather than answers from a chat interface alone. Access remains tied to Max, with Pro access planned for later.3 Before using it for business workflows, companies should decide which local files and apps an agent may access and which actions need review before completion.
What it means for companies
If you test the agent for business workflows, start with a narrowly defined task and review its access to files and apps. Check drafts and actions before messages are sent or files are changed.
Tencent releases HY-World 2.0 for interactive 3D worlds
The open-source project accepts text, images, and video. Code and model weights are initially available for part of the system.
Tencent introduced HY-World 2.0 as an open-source project in mid-April.1 The multimodal system is designed to generate, reconstruct, and simulate interactive 3D worlds from text, images, and video.12 The release is staged: code and model weights are available for part of the system, rather than every announced component.34
Unlike video-first approaches, HY-World 2.0 produces spatial 3D representations intended for use in game engines and embodied simulation workflows.12 That distinction matters to game developers and robotics teams that need 3D assets for downstream tools, not just rendered footage.15 Tencent describes interactive simulation, but the available materials do not establish measured physical accuracy against a dedicated physics benchmark.12
What it means for companies
If you build 3D environments for games or robotics, test the available components in your existing workflows. Check geometry and physical behavior yourself before using generated worlds in production.
Sources (5)
- 1 We're open-sourcing HY-World 2.0, a multimodal world model ... x.com
- 2 HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds arxiv.org
- 3 tencent/HY-World-2.0 huggingface.co
- 4 README.md · tencent/HY-World-2.0 at ... huggingface.co
- 5 Create README.md · tencent/HY-World-2.0 at 5c649d1 huggingface.co
Physical Intelligence shows robots learning from verbal coaching
Research model π0.7 completed an unfamiliar task with step-by-step instructions. Its reliability beyond the demonstrations remains unclear.
Physical Intelligence introduced its π0.7 robotics model in a research post on April 16, 2026.1 The model is intended to perform dexterous tasks and follow new language instructions, including for tasks absent from its training data.1 In a demonstration, step-by-step coaching helped a robot complete an unfamiliar task involving a kitchen appliance.12 Without that coaching, it made only a passable attempt, according to TechCrunch.2
The model builds on the company’s earlier π0 research.1 The demonstration suggests that language could guide robots through new procedures without collecting additional teleoperation data for every task.13 Physical Intelligence describes the results as early signs of generalization, however.1 Robust benchmark results, commercial availability details, and evidence of reliable performance beyond the demonstrated tasks are not available.12
What it means for companies
If you evaluate robots for changing workflows, test whether verbal instructions actually reduce the effort of setting up new tasks. Continue to assess reliability and failure cases in your own environment.
White House prepares Mythos access for US agencies
US agencies could receive controlled access to Anthropic’s Mythos. A firm authorization or timeline had not been confirmed.
The White House was preparing in mid-April to make a version of Anthropic’s Mythos available to federal agencies. The Office of Management and Budget was working on safeguards to let agencies use the model. But an email from federal CIO Gregory Barbaccia gave neither a firm commitment nor a timeline. It remained unclear whether the planned access had gone live.1
The plan comes amid a dispute between Anthropic and the Defense Department. The Pentagon had designated the company a supply-chain risk, barring use by the department and its contractors.23 Some federal agencies were nevertheless already testing the model, according to reports.2 Anthropic is providing Mythos through a restricted cybersecurity-focused deployment and had not planned a public release.4 Agency testing should therefore not be mistaken for an authorized, routine rollout.12
What it means for companies
If you use AI for government or regulated clients, check approvals and procurement rules before testing. Keep an alternative available while Mythos access and usage terms remain unsettled.
Sources (4)
- 1 White House to give US agencies Anthropic Mythos access ... reuters.com
- 2 Federal agencies skirt Trump's Anthropic ban to test its advanced AI model, Politico reports reuters.com
- 3 Anthropic talking to the Trump administration about its next AI model, co-founder says reuters.com
- 4 White House and Anthropic CEO discuss working together ... - Reuters reuters.com
Anthropic explains context management and subagents in Claude Code
Anthropic recommends using subagents and managing context deliberately during long Claude Code tasks.
Anthropic has published guidance on managing long sessions in Claude Code.1 An April 20 post covers when to use subagents, while a separate guide explains session management.12 Anthropic recommends subagents for tasks whose intermediate output is not needed in the main conversation. Each works in a fresh context window and returns a synthesized result to the parent session.1
The guidance addresses “context rot,” Anthropic’s term for performance degradation as a session’s context grows.1 Delegating a discrete task can keep unnecessary output out of the main context. Anthropic also describes compacting or clearing context to manage longer tasks.1 These are usage recommendations, not evidence of a new model release or measured performance gains from the approach.12
What it means for companies
If you use Claude Code for long development tasks, delegate self-contained work to subagents when the main session does not need the intermediate output. Before continuing, check whether older context is still useful or should be compacted.
Which of these developments matters for your company?
We help you turn AI news into concrete use cases, from assessment to implementation.
Book a free consultationEvery week we analyze a wide range of AI sources, select the stories that matter most to companies and research each of them. The texts are written with AI assistance and link to the original sources. How our news agent works