Tag: Agentic AI

  • Meta Launches Muse Spark 1.1: A New Frontier Agentic Model Enters the Paid API Market

    Meta Launches Muse Spark 1.1: A New Frontier Agentic Model Enters the Paid API Market

    Meta Superintelligence Labs released Muse Spark 1.1 on July 9, 2026, a multimodal reasoning model built specifically for agentic tasks that marks a significant strategic shift for the company. For the first time, Meta is charging for access to a frontier AI model through the paid Meta Model API, putting it in direct competition with Anthropic’s Claude and OpenAI’s GPT lineup. The launch was punctuated by CEO Mark Zuckerberg’s return to X after three years away from the platform. Muse Spark 1.1 arrives with a 1 million token context window, native computer use capabilities, and parallel sub-agent execution, entering public preview immediately for developers globally.

    What Was Announced

    Muse Spark 1.1 was released by Meta Superintelligence Labs, the research division led by Alexandr Wang, on July 9, 2026. The model is designed to handle complex, multi-step agentic workflows — a class of AI task that requires reasoning over long sessions, executing actions across computer interfaces, and managing many subtasks in parallel.

    Pricing for Muse Spark 1.1 is set at $1.25 per million input tokens and $4.25 per million output tokens. Developers can begin testing immediately with $20 in free API credits. The model is available through the Meta Model API in public preview, and is also accessible through the Meta AI app’s Thinking mode and at meta.ai, giving both enterprise developers and individual users access to the same underlying capability.

    CEO Mark Zuckerberg announced the launch on X, marking his return to the platform for the first time in three years — his last engagement there was in July 2023, when the platform rebranded from Twitter. Zuckerberg described Muse Spark 1.1 as “a strong agentic and coding model at a very low price,” signaling that Meta intends to compete on cost as well as raw capability.

    Alexandr Wang, who leads Meta Superintelligence Labs, said the new platform represents the company’s strongest model for agentic and coding work, with a focus on enabling autonomous multi-step task completion at enterprise scale.

    Technical Details

    Muse Spark 1.1 is built on a multimodal architecture trained for high performance on extended, multi-step tasks. The model supports a 1 million token context window, allowing it to retain information and reason across very long sessions without losing track of earlier context — an essential feature for enterprise workflows that may unfold over hours rather than minutes.

    One of the model’s key technical differentiators is its approach to parallel execution. Rather than processing complex tasks sequentially, Muse Spark 1.1 is trained to spawn and coordinate parallel sub-agents, enabling it to complete more steps in less time on large projects. The model also ships with native computer use capabilities, allowing it to interact directly with desktop applications, mobile interfaces, and web browsers to complete multi-step digital workflows autonomously.

    On benchmark evaluations, Muse Spark 1.1 tops professional and scaled tool-use benchmarks including JobBench and MCP Atlas. Meta reports major improvements over the original Muse Spark across tool use, computer use, coding, and multi-agent orchestration. The model trails Anthropic’s Opus 4.8 and OpenAI’s GPT-5.5 on pure coding and multimodal reasoning tasks, pointing to clear strengths in agentic and workflow automation scenarios.

    Industry Impact and Reactions

    The most significant aspect of the Muse Spark 1.1 release may not be the model itself, but what it signals about Meta’s business strategy. For years, Meta positioned itself as a champion of open-source AI, releasing its LLaMA model family freely and building a public reputation in contrast to closed API providers like Anthropic and OpenAI. The launch of a paid Meta Model API changes that equation directly. Meta is now entering the commercial frontier model market, offering a product that competes on price, capability, and a distinct technical focus on agentic tasks.

    The timing of the launch is notable. The AI coding and agentic AI markets have been intensifying rapidly throughout 2026, with major releases from virtually every large AI lab. Meta’s entry into this space with a model specifically designed for agentic and tool-use tasks puts additional pressure on the pricing tiers that Anthropic and OpenAI have established. At $1.25 per million input tokens, Muse Spark 1.1 is positioned as a cost-competitive option for developers building applications that make heavy use of AI tool calls and computer use.

    The fact that Zuckerberg personally returned to X to make the announcement underscores how significant Meta views this launch internally. The three-year absence from the platform made the post immediately visible to tech media and the developer community, amplifying the announcement beyond what a standard press release would achieve.

    What Comes Next

    Meta has indicated that Muse Spark 1.1 is the beginning of a new product line rather than a standalone model release. The Meta Model API is launching in public preview, suggesting the company plans to expand availability, add enterprise-grade features such as private deployment and usage analytics, and iterate on the model rapidly in the months ahead. Developers can expect additional SDK support, expanded documentation, and broader regional availability as the preview progresses.

    The competitive landscape will almost certainly respond. Anthropic, OpenAI, and Google have each made significant investments in agentic AI capabilities throughout 2026, and Meta’s entry at an aggressive price point adds further urgency to their own development roadmaps. The next benchmark releases from all four labs will be closely watched by enterprise buyers weighing platform commitments.

    Conclusion

    Meta Muse Spark 1.1 marks a meaningful turning point for the company and for the AI industry. A company long associated with open-source AI is now competing directly in the paid frontier model market, with a model purpose-built for agentic workflows, computer use, and large-scale task automation. Whether Muse Spark closes the performance gap with top competitors on coding and multimodal tasks in future versions remains to be seen, but the commercial and strategic implications of this launch extend well beyond any single benchmark result.

    Stay updated on the latest AI news at Evolve Digital.

  • Google Retires Gemini CLI: Antigravity CLI Takes Over as Google’s Premier AI Developer Platform

    Google Retires Gemini CLI: Antigravity CLI Takes Over as Google’s Premier AI Developer Platform

    Google officially retired its Gemini CLI developer tool on June 18, 2026, directing consumer and Google AI Pro and Ultra users to its new Antigravity CLI platform. The transition marks a significant shift in Google’s AI developer tooling strategy, moving from the open-source Gemini CLI — which had amassed over 100,000 GitHub stars — to a unified, closed-source agentic platform built for the next generation of AI-assisted software development. For the millions of developers who built automated workflows and CI/CD pipelines around Gemini CLI, today’s sunset is both an end and a beginning.

    What Was Announced

    On May 19, 2026, Google product managers Dmitry Lyalin and Taylor Mullen published an announcement on the Google Developers Blog confirming that Gemini CLI and Gemini Code Assist IDE extensions would cease serving requests for Google AI Pro and Ultra users on June 18, 2026. The post acknowledged the product’s remarkable open-source run, noting that Gemini CLI had achieved “over 100,000 GitHub stars, 6,000 merged pull requests, and hundreds of contributors” since its launch.

    The replacement platform is Antigravity CLI, invoked via the agy binary, which is built in Go and designed around an asynchronous, agent-first architecture. It shares the same underlying harness as the Antigravity 2.0 desktop application, creating a unified developer experience across terminal and graphical environments. Google is positioning Antigravity as its premier agentic development platform, consolidating developer-facing AI tools under a single brand.

    Enterprise customers with paid Gemini Code Assist Standard or Enterprise licenses, or those accessing Gemini models via paid API keys, retain uninterrupted access to the legacy Gemini CLI. Google also confirmed that GitHub organization users of Gemini Code Assist for GitHub are unaffected by today’s consumer-side retirement.

    Consumer users and Google AI Pro and Ultra subscribers who have not yet migrated lost access to Gemini CLI authentication as of today, June 18, 2026. Migration documentation is available immediately through Google’s Antigravity developer portal, with video walkthroughs scheduled for release in the coming weeks.

    Technical Details

    Antigravity CLI introduces several meaningful technical improvements over Gemini CLI. The most fundamental change is the shift to asynchronous agent orchestration. Where Gemini CLI blocked the terminal during complex or long-running tasks, Antigravity CLI can coordinate multiple background agents simultaneously. This allows developers to initiate large-scale code refactors, multi-step research tasks, or extended automated workflows without locking up their primary terminal session.

    The binary itself is written in Go, replacing the TypeScript foundation of the original Gemini CLI. This results in faster startup times and more responsive execution across terminal environments. All of the core developer-facing capabilities from Gemini CLI have been preserved and migrated to the Antigravity platform: Agent Skills carry over without modification, Hooks are fully supported, Subagents continue to function, and Extensions have been renamed Plugins under the new naming convention.

    The compute quota model has also been redesigned. Gemini CLI operated on a 1,000 requests-per-day cap, a structure suited to brief, discrete interactions. Antigravity CLI shifts to a weekly compute-based quota, better accommodating the more resource-intensive, long-running agentic tasks that the new async architecture is designed to handle. Developers with complex automated pipelines should review the new quota documentation to assess any impact on their workflows.

    Industry Impact and Reactions

    Google’s transition from Gemini CLI to Antigravity reflects a broader strategic pivot happening across the AI tooling industry. The move from conversational, request-response AI interfaces toward persistent, autonomous agentic platforms is accelerating at all major AI companies. Anthropic’s Claude Code and OpenAI’s Codex have similarly evolved into full development agents capable of controlling compute environments, managing files, and executing multi-step automated workflows.

    For Google specifically, the consolidation under the Antigravity brand is strategically significant. By unifying the terminal CLI and the desktop application under a shared agent harness, Google is positioning itself to compete directly with integrated agentic development environments rather than remaining a provider of standalone AI tools. This mirrors Anthropic’s approach with Claude Code, which runs the same agent runtime across CLI, desktop, and IDE extension contexts.

    The forced migration has drawn mixed reactions from the developer community. Performance improvements and the new async capabilities have been broadly welcomed, but the closure of Gemini CLI’s open-source repository in favor of a closed-source Go binary has drawn criticism. The Gemini CLI’s 6,000 merged pull requests represented a significant community investment, and the shift to a proprietary platform means that community contribution pathway closes with today’s retirement.

    What Comes Next

    Google has confirmed that all future model improvements and new agentic features will be delivered exclusively through the Antigravity platform. Enterprise customers currently on legacy Gemini CLI access will face the same migration choice over time, as the Antigravity ecosystem becomes the primary vehicle for accessing Google’s frontier AI models in developer contexts. For most developers, the practical timeline for migration is now: consumer accounts have already lost access, and Google’s roadmap signals Antigravity as the sole long-term path.

    Migration documentation is live as of today, with full video walkthroughs releasing in the coming weeks to guide developers through the transition from Gemini CLI workflows to their Antigravity equivalents. Developers are advised to audit any existing CI/CD pipelines, scripts, or automations that reference the gemini command and plan their migration to the agy binary accordingly before any dependent systems experience disruption.

    Conclusion

    The Gemini CLI sunset on June 18, 2026 closes the book on one of the most successful open-source AI developer tools of the past two years. With Antigravity CLI now at the center of Google’s developer AI strategy, the company is making a clear bet on asynchronous, agent-first tooling as the foundation of modern software development workflows. The transition reflects an industry-wide shift: the era of interactive chat-style AI assistants is giving way to persistent, autonomous agentic platforms that can operate independently across complex, multi-step tasks. Developers who migrate quickly will be best positioned to take advantage of the capabilities that Antigravity’s unified architecture makes possible.

    Stay updated on the latest AI news at Evolve Digital.

  • Microsoft Build 2026: Windows Gains On-Device Aion AI Models, Copilot Runtime, and Agentic Tools

    Microsoft Build 2026: Windows Gains On-Device Aion AI Models, Copilot Runtime, and Agentic Tools

    Microsoft opened its annual Build developer conference on June 2, 2026, with a keynote led by CEO Satya Nadella that placed artificial intelligence at the center of the Windows platform strategy. The event, held at Fort Mason Center in San Francisco and streamed globally, delivered a significant range of AI announcements targeting developers, enterprises, and end users. From new on-device language models shipping inside Windows to enterprise-grade agent governance tools, Build 2026 marks one of the most AI-dense Microsoft developer events in recent memory.

    What Was Announced

    The headline product for developers is Aion 1.0, a new family of small language models (SLMs) built by Microsoft specifically for on-device Windows workloads. Two variants were previewed: Aion 1.0 Instruct, a compact model optimized for everyday text intelligence tasks including summarization, rewrites, intent recognition, and accessibility features; and Aion 1.0 Plan, a 14-billion-parameter reasoning and tool-calling model with a 32K context window that will ship in-box with Windows.

    Alongside the Aion models, Microsoft unveiled Copilot Runtime for Windows, a suite of local inference APIs that allow Win32 and WinUI 3 applications to tap into the same on-device AI models that power the operating system’s Copilot experience. This means developers can build Windows applications that perform AI tasks locally, without sending data to the cloud. Windows AI APIs are also being extended beyond Copilot+ PC hardware to support GPU acceleration for Phi Silica and CPU-based execution for video super resolution and live captions.

    A new Speech Recognition API, now in preview, delivers real-time on-device speech-to-text from any audio source, including microphone, stream, or file, with hardware-accelerated execution on CPU or NPU. This capability opens new opportunities for developers building transcription, accessibility, and voice-driven applications for Windows.

    On the infrastructure side, Microsoft announced Azure Agent Mesh, a new service designed to orchestrate AI agents that span multiple cloud environments, on-premises systems, and edge devices, enabling large organizations to build and manage heterogeneous multi-agent systems at scale.

    Technical Details

    The Aion 1.0 Plan model’s 14-billion-parameter scale and 32K context length place it in a competitive range for local reasoning tasks. Shipping the model in-box with Windows removes the installation and configuration barrier that has historically limited on-device AI adoption. Microsoft’s Copilot Runtime abstracts hardware differences, presenting a unified API surface regardless of whether the underlying execution is on NPU, GPU, or CPU, a significant engineering decision that broadens the range of Windows hardware capable of running AI-accelerated applications natively.

    AgentGuard, Microsoft’s new enterprise governance layer for AI agents, enforces role-based access permissions, data loss prevention policies, and comprehensive audit logging across all agent interactions. The capability is designed to address enterprise compliance and security requirements as organizations deploy autonomous AI agents across their workflows. AgentGuard integrates directly with Microsoft’s existing identity and compliance tooling.

    The Surface RTX Spark Dev Box, announced alongside the software stack, is a compact developer workstation powered by an NVIDIA RTX Spark module with 1 petaflop of AI compute and 128 GB of unified memory. It is capable of running models up to 120 billion parameters locally, giving developers a self-contained environment for building and testing large model applications without cloud dependency.

    Industry Impact and Reactions

    Microsoft’s Build 2026 announcements represent a strategic push to make Windows the primary platform for AI-native application development. By shipping Aion 1.0 models in-box and providing Copilot Runtime APIs, Microsoft is positioning the operating system itself as an AI infrastructure layer, a significant shift from the traditional view of Windows as a software delivery platform. This approach competes directly with cloud-first AI strategies by bringing inference capability directly to the device.

    The Azure Agent Mesh announcement signals Microsoft’s intent to capture enterprise demand for multi-agent AI orchestration at scale. With organizations increasingly deploying AI agents across business processes, a managed cross-cloud orchestration service addresses a real operational gap. The addition of AgentGuard’s compliance and governance capabilities shows Microsoft is addressing enterprise risk concerns that have slowed AI agent adoption in regulated industries.

    The Surface RTX Spark Dev Box underscores the broader trend of purpose-built AI developer hardware. By pairing high-memory NVIDIA RTX Spark silicon with 128 GB of unified memory, Microsoft is offering developers a machine that can run very large models locally, reducing the latency and cost associated with cloud-based development and testing cycles.

    What Comes Next

    Microsoft Build 2026 continues through June 3, with additional sessions and developer workshops expected to provide deeper technical detail on Aion 1.0, Copilot Runtime APIs, and Azure Agent Mesh. The Aion 1.0 Instruct and Plan models are currently in preview, with general availability timelines not yet confirmed. Developers interested in early access can register through the Windows AI developer program.

    Broader Windows rollout for the new AI APIs and in-box Aion model support is anticipated to follow through future Windows Update releases, though Microsoft has not confirmed a specific date. Enterprise customers interested in AgentGuard and Azure Agent Mesh can explore preview enrollment through the Azure portal.

    Conclusion

    Microsoft Build 2026 delivers one of the most comprehensive AI platform updates in the company’s developer conference history. The combination of on-device Aion models shipping in Windows, Copilot Runtime APIs for app developers, cross-cloud agent orchestration through Azure Agent Mesh, and the governance controls in AgentGuard paints a detailed picture of Microsoft’s strategy: make every Windows device an AI-capable endpoint and make Azure the management plane for enterprise AI agents at scale. The announcements confirm that the operating system itself is becoming an active participant in the AI application stack.

    Stay updated on the latest AI news at Evolve Digital.

  • Anthropic Releases Claude Opus 4.8 With Dynamic Workflows and Major Coding Improvements

    Anthropic Releases Claude Opus 4.8 With Dynamic Workflows and Major Coding Improvements

    Anthropic has released Claude Opus 4.8, the latest iteration of its flagship AI model, bringing meaningful gains in coding reliability, reasoning, and autonomous operation. Released on May 29, 2026, just 41 days after Opus 4.7, the update introduces a headline new capability called Dynamic Workflows and delivers measurable benchmark improvements across core performance areas. The model is available globally today via the Anthropic API and Claude.ai at the same price point as its predecessor.

    What Was Announced

    Anthropic described Claude Opus 4.8 as offering “sharper judgment, more honesty about its progress, and the ability to work independently for longer than its predecessors.” The company released benchmark data showing improvements on two key metrics: agentic coding performance rose from 64.3% to 69.2%, while multidisciplinary reasoning with tools improved from 54.7% to 57.9%.

    One of the more notable reliability improvements is in code quality oversight. Anthropic says Opus 4.8 is approximately four times less likely than Opus 4.7 to allow flaws in code it has written to pass silently without flagging them, addressing a persistent pain point for teams relying on AI models in software development pipelines.

    Speed also improved: the Opus 4.8 fast mode is roughly 2.5 times quicker than the equivalent mode in Opus 4.7. Critically, Anthropic kept pricing identical to the previous model version, meaning existing API users receive the full upgrade at no additional cost.

    The centerpiece of the release is Dynamic Workflows, now available in research preview. This feature is designed to enable Opus 4.8 to coordinate and manage complex, long-horizon tasks by orchestrating hundreds of parallel subagents simultaneously. Anthropic positioned this capability specifically for enterprise teams building large-scale agentic pipelines where multiple AI instances must collaborate on a shared goal.

    Technical Details

    Dynamic Workflows represents a significant architectural extension of how Claude operates in multi-agent contexts. Rather than functioning as a single model responding sequentially, Opus 4.8 with Dynamic Workflows acts as an orchestrator, delegating subtasks to parallel subagents and synthesizing their outputs into coherent results. This allows the model to tackle problems that would be impractical to complete within a single context window or within the latency constraints of a linear workflow.

    The coding improvements in Opus 4.8 are tied closely to enhancements in self-monitoring. The model shows improved ability to recognize when its own output contains errors or uncertainties, and to flag these rather than proceeding with flawed assumptions. This behavioral shift is particularly significant in autonomous coding scenarios, where silent errors can propagate through large codebases before being detected.

    Anthropic also notes that fast mode throughput improvements were achieved through inference optimizations rather than model compression, preserving the underlying capability profile of the model while significantly reducing latency for time-sensitive applications.

    Industry Impact and Reactions

    The release comes in a period of rapid iteration across the frontier AI model landscape. Anthropic’s 41-day release cycle from Opus 4.7 to 4.8 signals a faster cadence than the company has historically maintained, reflecting competitive pressure from OpenAI and Google, both of which have accelerated their own release timelines in 2026.

    The combination of Dynamic Workflows and improved coding reliability is directly relevant to the growing enterprise market for agentic AI. Businesses deploying AI in software development, data analysis, and automated workflow management stand to benefit most from the improvements. The fact that the upgrade carries no price increase removes one of the traditional adoption barriers for enterprise customers already on the Anthropic API.

    Claude Opus 4.8 also arrives alongside a significant financial milestone for Anthropic: the company recently raised additional private funding, reaching a valuation of approximately $965 billion. This financial backdrop gives Anthropic substantial runway to continue research investment and infrastructure expansion as it competes at the frontier of large language model development.

    What Comes Next

    Dynamic Workflows is currently in research preview, suggesting Anthropic is gathering feedback before a broader production release. The company has not announced a specific general availability date for the feature, but the research preview designation typically precedes a full rollout within weeks to months. Anthropic is also expected to bring its next class of models, which the company has referred to informally as Mythos-class, to a wider set of customers later in 2026.

    For teams already using Opus 4.7, the path to Opus 4.8 requires only updating to the latest model version in the API — no integration changes are needed to access the core improvements. Teams interested in Dynamic Workflows will need to apply for the research preview through Anthropic’s developer portal.

    Conclusion

    Claude Opus 4.8 represents a focused, evidence-based upgrade to one of the leading frontier AI models currently available. With improved coding reliability, faster inference, and the introduction of Dynamic Workflows, Anthropic is addressing the real-world needs of developers and enterprises building agentic AI systems. The decision to maintain existing pricing makes this a straightforward upgrade for current users, and positions Anthropic competitively as the race to deploy capable, reliable AI agents in enterprise environments continues to intensify.

    Stay updated on the latest AI news at Evolve Digital.