Agent analysis
Standing syntheses the agent writes over each lab's captured pages, structured signals, and bounded web evidence — every material claim cited back to its source.
DeepSeekDeepSeek3wDeepSeek is executing a two-pronged strategy in mid-2026: aggressively optimizing inference economics through open-source speculative decoding infrastructure (DSpark, DeepSpec, Eagle3) while simultaneously expanding into the agentic coding product market with a new Code Harness team. The lab's release cadence shows a shift from pure model releases (R1, V3 series) toward inference-serving tooling that reduces…CohereCohere3wCohere is transitioning from a research-forward multilingual lab into a security-first enterprise platform company. The evidence shows a lab simultaneously pushing research on MoE architectures and synthetic data while executing an aggressive commercialization buildout centered on the North platform. The North Mini Code release — a 30B MoE model with 3B active parameters for agentic coding under Apache 2.0 — is the…Meituan (LongCat)Meituan (LongCat)3wMeituan's LongCat lab is executing a full-spectrum, model-system co-design strategy that spans text, image, video, audio, omni-modal generation, formal reasoning, and agentic evaluation — all anchored by a growing family of open-weight Mixture-of-Experts models. The capstone is LongCat-2.0, a 1.6T-parameter MoE with ~48B activated parameters trained and deployed end-to-end on domestic Chinese AI ASIC superpods,…LG AI Research (EXAONE)LG AI Research (EXAONE)3wLG AI Research is executing a multi-vector expansion strategy anchored on its proprietary EXAONE model family, with three clear thrusts emerging from the 2025–2026 evidence: (1) a bold pivot into Physical AI and Robot Foundation Models (RFMs) built atop the EXAONE foundation stack, (2) deep commercialization via enterprise BD, defense-sector AI, and customer success hires, and (3) a structured model-release cadence…InclusionAI (Ant Group)InclusionAI (Ant Group)3wInclusionAI operates as Ant Group's open-source research and release vehicle, pursuing a dual-track strategy: (1) high-capacity MoE foundation models (Ling/Ring families up to 1T parameters) with hybrid linear attention for efficient long-context serving, and (2) a growing stack of agentic infrastructure — RL post-training tooling (AReno), multi-agent runtimes (AWorld), and multimodal safety guardrails (SingGuard).…CoreWeaveCoreWeave3wCoreWeave is transitioning from a GPU-rental neocloud into a full-stack AI cloud platform with serious public-company scale and discipline. The evidence points to three reinforcing vectors: (1) infrastructure velocity — $5B annual revenue growing 168% YoY, 850MW active power across 43 data centers globally, and first-to-market NVIDIA Vera Rubin NVL72 bring-up; (2) software-stack deepening — a unified agentic AI…Cloudflare (Workers AI)Cloudflare (Workers AI)3wCloudflare is executing a multi-vector pivot from content-delivery infrastructure toward AI-native compute and agentic intermediation. The evidence shows three reinforcing moves: (1) Workers AI is being hardened as an inference platform for trillion-parameter open-weight models with commercial-grade routing and billing, (2) an Agents SDK and first-party agent framework (Flue) are being layered on top to capture…ByteDance (Doubao/Seed)ByteDance (Doubao/Seed)3wByteDance Seed is executing a deliberate multi-frontier strategy: shipping production models through Volcano Engine (Doubao 2.1 Pro, Seedance 2.5) while simultaneously open-sourcing a broad portfolio of research artifacts across LLM reasoning, multimodal understanding, video, biology, agents, and systems infrastructure. The evidence reveals a lab that is scaling inference infrastructure aggressively (VeOmni,…AnthropicAnthropic3wAnthropic is transitioning from a model-research lab into a vertically integrated AI product company with global commercial ambitions. The evidence shows simultaneous acceleration across four fronts: (1) new model launches (Sonnet 5, Fable 5, Mythos 5) alongside the lifting of US export controls [W1, W3, E2]; (2) a buildout of vertical AI products for financial services, life sciences, and scientific research —…IBM (Granite)IBM (Granite)3w{"content": "## Thesis\n\nIBM is converging its Granite open-source model family with a multi-billion-dollar infrastructure and security play to position as the enterprise-trusted AI platform. The lab ships Apache 2.0-licensed models across seven modalities — language, code, vision, speech, time series, embeddings, and geospatial — while simultaneously building the hardware substrate (sub-1nm chips) and supply-chain…BasetenBaseten3wBaseten is in the midst of a breakout scaling phase fueled by a $1.5B Series F. The company is tripling headcount while simultaneously shipping at high velocity across three vectors: developer tooling (CLI, MCP server, Chains GA), model performance research (BEI embeddings, speculative decoding, timestep distillation), and GPU infrastructure breadth (B200, GH200, H100/H200, multi-node). The hiring pattern skews…Amazon (Nova)Amazon (Nova)3wAmazon is building a vertically integrated AI stack that runs from custom silicon and formally verified cloud infrastructure through foundation models, agent frameworks, and domain-specific evaluation tooling. The evidence reveals a three-pillar strategy: (1) infrastructure differentiation via Graviton5 chiplet architecture and the formally verified Nitro Isolation Engine, (2) a multi-modal foundation model…ClarifaiClarifai3wClarifai is in wind-down mode as a standalone entity. The company's platform and inference IP have been acquired by Nebius (NASDAQ: NBIS), with all Clarifai services ceasing on July 17, 2026. New signups closed May 19 and new payments stop June 17, 2026. Recent GitHub activity reflects a final synchronization push of gRPC SDKs across seven languages to version 12.5.1, with no release notes published—consistent with…Inception LabsInception Labs3wInception Labs is building a new architectural wedge into the LLM market by replacing sequential autoregressive decoding with parallel diffusion-based generation. Founded by Stanford professor Stefano Ermon alongside Aditya Grover (CTO) and Volodymyr Kuleshov, the company launched the Mercury family — the first commercial-scale diffusion large language models (dLLMs) — in February 2025 and released Mercury 2 as its…AI21 LabsAI21 Labs3wAI21 Labs is executing a sharp, survival-driven pivot: it is ceasing to compete as a frontier model provider and is instead concentrating entirely on enterprise AI agent orchestration via its Maestro platform. The evidence shows a lab that has recognized it cannot win the foundation-model arms race and is retreating to a defensible niche — reliable, auditable, "boring" agentic systems for regulated industries. The…CerebrasCerebras3wCerebras is executing a hard pivot from wafer-scale training hardware specialist to full-stack inference cloud provider. The evidence shows a company that IPO'd in May 2026, closed an $850M revolving credit facility, and is aggressively building out inference datacenters across North America and Europe with a target of 20x aggregate capacity expansion. The wafer-scale architecture delivers inference speeds 10–20x…Databricks (DBRX)Databricks (DBRX)4wDatabricks is executing a three-front offensive: it is monetising enterprise data gravity through Lakebase (a serverless Postgres for operational/AI workloads), hardening an agent platform (Genie One, Genie Agents, Omnigent) that turns the Lakehouse into an operating system for enterprise agents, and building a specialised research pipeline — data agents trained with RL — that seeks to match frontier model…Arcee AIArcee AI4wArcee AI is executing a deliberate transition from a post-training and SLM-adaptation specialist into a full-stack frontier AI lab. The evidence maps this in two phases: a tooling-and-distillation era (MergeKit, DistillKit, SuperNova family) spanning 2023–mid-2025, followed by a from-scratch pretraining era (AFM-4.5B, Trinity MoE family) beginning mid-2025 and accelerating into 2026. The lab operates with striking…StepFunStepFun4wStepFun is a Shanghai-based frontier AI lab executing an unusually broad multimodal strategy—spanning video generation, real-time voice, image editing, 3D asset generation, formal mathematics, GUI agents, and deep research—while releasing the vast majority of its model weights, code, and benchmarks under permissive open-source licenses. The lab's May 2026 release of Step 3.7 Flash, a 198B MoE vision-language model…Xiaomi (MiMo)Xiaomi (MiMo)4wXiaomiMiMo is executing a full-stack, open-source AI strategy that spans text reasoning, vision, audio, embodied AI, and coding agents — with an accelerating pivot toward the Agent era in mid-2026. The lab's evidence trail reveals a deliberate arc: a reasoning-first 7B model family born from pretraining-to-posttraining optimization; rapid horizontal expansion into vision-language (MiMo-VL, May 2025), audio language…Sarvam AISarvam AI4wSarvam AI is executing a three-horizon transition from sovereign AI R&D lab to full-stack platform company competing globally. The evidence pack captures this inflection: a unicorn-level fundraise ($300M at ~$1.5B valuation), the March 2026 release of two MoE reasoning models — Sarvam-30B (32B params) and Sarvam-105B (106B params) — trained from scratch on IndiaAI Mission compute, and a hiring wave of ~50+ open…OpenBMB (MiniCPM)OpenBMB (MiniCPM)4wOpenBMB is a university-tethered research-to-product organization—with author affiliations spanning Tsinghua University and Northeastern University —pursuing a clear thesis: compact, on-device frontier AI that ships openly and runs locally. Their portfolio spans language models (MiniCPM series, CPM-Bee), vision-language models (MiniCPM-V, VisCPM), speech synthesis (VoxCPM), agent frameworks (AgentCPM, XAgent,…Upstage (Solar)Upstage (Solar)4wUpstage is executing a platform consolidation play, moving from a pure models-and-APIs company toward an integrated AI stack combining proprietary Solar LLMs, acquired portal (Daum) assets, and agentic platforms (Timely). Core thesis: Upstage is building an AI-for-everyone ecosystem — not just enterprise AI, but consumer-facing AI through a portal. The Solar model family (10.7B→31B→Open 100B→22B in preview/pro)…Reka AIReka AI4wReka is evolving from a multimodal model builder into a physical-AI company. The evidence shows a lab that ships compact, deployment-flexible vision-language models (Flash at 21B, Edge at 7B), layers enterprise agentic platforms on top (Nexus, Vision, Research), and is now orienting research toward world models, embodied data, and physical reasoning. The $110M raise backed by NVIDIA and Snowflake, the Moonvalley…Nous ResearchNous Research4wNous Research is in the midst of a rapid productization pivot. Hermes Agent — a 190k-star open-source AI agent platform — is receiving weekly releases with ~1,475 commits per cycle and 245 community contributors. Simultaneously, the org is shipping frontier models (Hermes-4 family through 405B scale), building RL training infrastructure (Atropos at 1,273 stars), and advancing distributed training research…Eigen AIEigen AI4wEigen AI is a 2025-founded inference optimization company acquired by NASDAQ-listed neocloud Nebius for $643M, with the deal closing on 10 June 2026. The lab's optimization stack is being integrated into Nebius Token Factory to deliver production inference at scale. Active hiring for post-training/inference engineering and platform product management signals a dual buildout: continued technical R&D on model serving…WaferWafer4wWafer is a hardware-centric AI inference platform building competitive advantage through GPU kernel optimization expertise, with a distinctive multi-vendor strategy spanning NVIDIA and AMD accelerators. The evidence depicts a company vertically integrated from low-level kernel engineering up to a serverless inference product, using public benchmarks and developer education content as both recruiting and go-to-market…Public AIPublic AI4wPublic AI is not a frontier model lab; it is a neocloud-adjacent public infrastructure play building an inference utility positioned as an open, democratically governed alternative to commercial AI APIs. The org's GitHub activity reveals a concentrated push to operationalize a chat-based inference platform (chat.publicai.co) atop OpenWebUI, with CI/CD pipeline maturation, API gateway deployment, and multi-geography…CompactifAI (Multiverse Computing)CompactifAI (Multiverse Computing)4wCompactifAI (Multiverse Computing) is transitioning from a quantum-software R&D shop into a commercial AI infrastructure company anchored by model compression. Its proprietary CompactifAI technology applies quantum-inspired tensor-network mathematics to prune and restructure pre-trained LLMs, producing "Slim" variants that retain reasoning and tool-use capabilities at reduced inference cost. The lab is running a…Blackbox AIBlackbox AI4wBlackbox AI is not a frontier model builder; it is an inference-infrastructure and agent-orchestration platform that competes on serving others' models faster, cheaper, and more securely than anyone else. The company's public signals converge on a single bet: that enterprise and government adoption of coding agents will be won at the orchestration and inference layer, not at the model-training layer [W1, W5]. With…MakoraMakora4wMakora is a performance-engineering organization focused on automated GPU kernel generation and inference optimization. Its public surface spans four categories: (1) an AI-driven kernel generation system (MakoraGenerate) that produces optimized GPU kernels targeting NVIDIA H100/B200, AMD MI300X, and Tenstorrent hardware; (2) a lightweight multi-vendor GPU querying utility (gpuq) supporting CUDA and HIP runtimes; (3)…StreamLake (Kuaishou)StreamLake (Kuaishou)4wKuaishou's StreamLake is executing a dual-pronged open-source strategy: a video-native multimodal foundation model line (Keye-VL-2.0) aimed at long-video understanding and agentic capabilities, and an AI coding product suite (KAT-Coder, CodeFlicker, Vanchin) targeting the software engineering tool market. Both tracks are anchored in Apache 2.0 releases, rigorous public benchmarking against frontier models (GPT-5,…ParasailParasail4wParasail is an early-stage AI infrastructure company (Series A, $32M raised, $160M valuation) building a serverless inference cloud that aggregates distributed GPU supply into an OpenAI-compatible platform for open-weight models. The company positions itself as a GPU-network orchestration layer: workloads are automatically matched across a multi-provider GPU network, freeing developers from vendor lock-in and…GMI CloudGMI Cloud4wGMI Cloud is an inference-optimized neocloud building a full-stack platform tightly coupled to NVIDIA's hardware roadmap. The evidence shows a company transitioning from bare-metal GPU provisioning to a managed platform layer: 10 open roles are clustered around a named "Inference Engine" product, AgentBox has shipped as an agent marketplace and hosting platform, and every public post ties GMI's infrastructure…ScalewayScaleway4wScaleway is executing a three-pillar strategy to differentiate as Europe's sovereign AI cloud: (1) serving frontier open-weight models through its Generative APIs platform as a managed alternative to proprietary hyperscalers, (2) embedding sustainability and CSRD compliance tooling directly into its cloud product portfolio, and (3) investing in European AI infrastructure sovereignty through consortia spanning…Novita AINovita AI4wNovita AI is executing a two-pronged evolution: it operates a commercial model API and agent sandbox platform for third-party frontier models, while simultaneously building deep inference infrastructure—most visibly pegaflow, a Rust-based KV cache storage engine with vLLM integration—that targets the performance bottleneck of large-scale LLM serving. The pattern of GTM hiring in San Mateo alongside a relentless…SiliconFlowSiliconFlow4wSiliconFlow is building a "Token Factory" — an AI inference infrastructure layer that normalizes heterogeneous compute into standardized token output. The GitHub evidence reveals a two-track product strategy: (1) a deep acceleration stack (OneDiff, Nexfort) that compiles diffusion and LLM workloads for faster inference, and (2) BizyAir, a cloud-hosted ComfyUI service that wraps model access into a managed developer…HyperbolicHyperbolic4wHyperbolic is in a post–Series A scaling sprint, pivoting from its early Web3/decentralized microservices roots into a full-stack GPU marketplace aggregator. The evidence shows a company simultaneously hiring for infrastructure depth (GPU orchestration, bare-metal provisioning, SRE), commercial operations (supply, finance, GTM), and developer tooling (CLI, MCP, AI SDK, Gradio). The recent Forge launch crystallizes…FriendliAIFriendliAI4wFriendliAI is an AI inference infrastructure company entering an aggressive commercialization phase, signaled by a $20M funding round, a rapid SDK iteration cadence with breaking API changes across all serving tiers, the launch of a public OpenAPI schema, and day-zero support for frontier open-weight models. The dual-hub (Seoul/San Francisco) hiring pattern reveals simultaneous investment in core inference engine…DigitalOcean (GradientAI)DigitalOcean (GradientAI)4wDigitalOcean (GradientAI) is executing a concentrated pivot into agentic AI infrastructure as a managed cloud service, building the full stack from GPU inference to hosted agent runtimes. The evidence reveals a coordinated three-pronged buildout: (1) an Inference Engine now generally available with frontier model support across OpenAI, Anthropic, and fal; (2) a Codex plugin in Public Preview that provisions…Lightning AILightning AI4wLightning AI is in the midst of a structural transformation from developer-framework shop into a vertically integrated neocloud. The merger with Voltage Park [P2, P3] has reshaped the company's operational DNA: it now owns and operates physical data centers across at least three US geographies (Quincy WA, Fort Worth TX, Lisle IL) [P22, P23], is building bare-metal GPU compute, storage, and observability…ReplicateReplicate4wReplicate is a post-acquisition platform operating as an inference API aggregator, not a model builder. Following its acquisition by Cloudflare, its activity centers on platform engineering — evidenced by an intense cog release cadence across v0.16–v0.21 — ecosystem integration (SDKs in Python and JavaScript, MCP, LangChain, agent skills), and positioning as the hosted API layer for third-party frontier and…DeepInfraDeepInfra4wDeepInfra is an inference-cloud provider exploiting the open-weight model boom, not a model-building lab. Its GitHub footprint reveals a company systematically forking and maintaining the full inference-serving stack — from CUDA kernels to serving engines to client SDKs — while its $107M Series B and targeted hiring confirm a bet on inference infrastructure as a standalone business. The org tracks frontier…Snowflake (Arctic)Snowflake (Arctic)4wSnowflake is executing a deliberate convergence play: its Arctic model family — specialized for SQL, code generation, and enterprise retrieval — is being positioned not as a standalone frontier contender but as the AI inference layer inside a governed, agentic data platform. The firm's public writing, hiring, and releases all orbit a single narrative: "the agentic enterprise". Arctic now spans speculators…SambaNova SystemsSambaNova Systems4wSambaNova Systems is executing a decisive pivot from AI training hardware toward becoming an inference cloud provider purpose-built for agentic AI workloads. The evidence pack captures a company compressing its stack around three interlocking bets: (1) disaggregated/hybrid inference pairing its own SN40 RDU with NVIDIA GPUs for prefill-decode splitting [E29, E53, W2]; (2) "premium inference" as a differentiated…Fireworks AIFireworks AI4wFireworks AI is a Series C ($4B valuation) generative AI infrastructure platform transitioning from inference-speed leader to full-stack AI cloud provider, with training, fine-tuning, serverless and dedicated inference, multi-LoRA serving, and agentic orchestration all built on proprietary infrastructure. The most recent evidence — spanning June 2026 — reveals a company in an intensive GTM buildout phase, anchored…GroqGroq4wGroq is rebuilding as a pure-play AI inference cloud after a transformative non-acquisition by Nvidia that took its founding CEO, president, and key engineers. A $650M raise in June 2026 aims to scale GroqCloud to 200MW by 2027 and serve 5M developers on its purpose-built LPU chip architecture. The evidence pack shows Groq rapidly maturing SDK tooling (Python v1.5.0, TypeScript v1.3.0), building an evaluation and…Together AITogether AI4wTogether AI is consolidating its positioning as the AI-native cloud — an inference-first infrastructure platform that competes on raw speed and cost per token. The evidence pack shows the company simultaneously building out in three directions: (1) deepening the infrastructure surface from GPU clusters into managed storage, networking, and observability, (2) layering enterprise trust and access-control primitives…NebiusNebius4wNebius is executing a multi-front AI cloud scaling thesis: it is simultaneously building out physical data center capacity across the US and Europe, expanding its GPU orchestration software stack, commercializing a new agentic search product (Tavily), and deepening its research bench through an acqui-hire (Clarifai). The hiring pattern reveals a company transitioning from infrastructure provider to full-stack AI…Baidu (ERNIE)Baidu (ERNIE)4wThesis: Baidu is on a dual track — shipping competitive frontier models (ERNIE 5.1 at claimed 6% training cost of peers, Unlimited-OCR, NAVA) while aggressively building US enterprise sales/GTM teams and a Sunnyvale silicon design org. The formation of a centralized Model Committee signals internal R&D consolidation as the company pivots toward the "agent era." Key signals:Google (DeepMind / Gemini)Google (DeepMind / Gemini)Jun 22Google DeepMind is executing a two-track strategy: shipping a rapid cadence of open-weight Gemma 4 models (Gemma 4 12B, DiffusionGemma, quantization-aware training variants) while simultaneously building the agentic and safety infrastructure for Gemini 3.5’s frontier deployment. The evidence reveals a lab investing heavily in agent safety frameworks, embodied reasoning, national-scale AI deployment partnerships, and…Zhipu AI (GLM)Zhipu AI (GLM)Jun 8Zhipu AI (GLM) is shipping a broad, fast-moving family of open-weight GLM models across text, vision, OCR, speech, and image generation, releasing point versions at high cadence (GLM-4.5 through GLM-5/5.1 plus specialized variants) and backing them with first-party SDKs. The standout signal is reach: its OCR and Flash models are pulling millions of monthly Hugging Face downloads, and its older ChatGLM line remains…xAIxAIJun 8xAI is in a productize-and-distribute phase: its frontier work (Grok) lives behind the API while the public footprint is dominated by developer tooling and open-weight artifacts of prior generations. The xai-sdk-python is shipping rapidly (five releases tracked, through v1.15.0), and the company has open-sourced both the Grok-1/Grok-2 weights and, notably, the X recommendation algorithm — signaling tight integration…Tencent HunyuanTencent HunyuanJun 8Tencent Hunyuan is running broad on open-weight generative media — its public footprint skews heavily toward 3D, video, image, and world-model generation rather than chat LLMs. The most-downloaded asset is a frontier-scale ~298B-param model (tencent/Hy3-preview, 90k downloads/30d), but the highest-starred surface area on GitHub is its visual-generation stack (Hunyuan3D, HunyuanVideo). It is also pushing into newer…Qwen (Alibaba Cloud)Qwen (Alibaba Cloud)Jun 8Qwen (Alibaba Cloud) is running one of the most prolific open-weight release cadences in the field, shipping a full ladder of dense and Mixture-of-Experts models — currently the Qwen3.5 and Qwen3.6 generations — across every modality and a parallel agentic coding stack (qwen-code, 25k stars). Adoption is enormous: its current flagship-tier checkpoints each pull millions of Hugging Face downloads in a 30-day window.…OpenAIOpenAIJun 8OpenAI is operating on two fronts at once: a frontier-model release cadence aimed at consumers and developers, and a hard pivot into agentic developer tooling. Its public footprint right now is dominated by Codex, a terminal coding agent shipping near-daily alpha builds, and a wave of GPT-5.x launches (GPT-5.5, GPT-5.4, GPT-5.3-Codex) that top Hacker News. The hiring and infra signals point to scaling compute and…NVIDIANVIDIAJun 8NVIDIA is positioning itself as the full-stack supplier of the "AI factory" era — selling not just silicon but open models, agent runtimes, and physical-AI foundation models that run on its hardware. The current push centers on three fronts: long-running agents (the Nemotron 3 Ultra family and the NemoClaw agent blueprint), physical/world AI (Cosmos 3 and robotics), and local/personal agents on new hardware (RTX…Moonshot AI (Kimi)Moonshot AI (Kimi)Jun 8Moonshot AI (Kimi) is shipping open-weight, trillion-parameter mixture-of-experts frontier models at a fast iteration cadence — the Kimi-K2 line is its flagship, now through K2.5 and K2.6 plus a dedicated K2-Thinking variant. Alongside the weights it is building a full agentic-coding surface (the kimi-cli / kimi-code tools) and publishing efficiency-oriented architecture research (linear attention, attention…Mistral AIMistral AIJun 8Mistral AI is executing a broad open-weights strategy across every modality and size tier at once: text instruct/reasoning models from 3B up to 128B, a Voxtral audio/speech family (realtime, TTS), and a Devstral coding line. Distribution runs through Hugging Face at serious volume and a full client/tooling stack (mistral-inference, mistral-common, multi-language SDKs). The 2512/2602/2603 release cadence shows rapid,…MiniMaxMiniMaxJun 8MiniMax is shipping fast-iterating, open-weight large language models — the MiniMax-M2 family dominates its current footprint, with the latest MiniMax-M2.7 (~229B params) already pulling 2.5M 30-day downloads on Hugging Face. Alongside the models it is building out the surrounding agent tooling (a CLI, an MCP server, an agent harness, and a published "skills" library), signaling a play to own not just the weights…MicrosoftMicrosoftJun 8Microsoft's public footprint is dominated by its developer-tools and platform empire — vscode (186k stars), PowerToys, TypeScript, terminal, and playwright — but the AI-specific signal sits in Microsoft Research, which is releasing small, domain-specific foundation models (materials, the electric grid) on Hugging Face rather than chasing a single frontier LLM. Recent shipping concentrates on agent infrastructure (an…Meta AI (Llama)Meta AI (Llama)Jun 8Meta AI is the open-weight anchor of the frontier-model field: it ships the Llama family under permissive licenses and lets the ecosystem do distribution, while pivoting its newest generation (Llama 4) to mixture-of-experts. Alongside the models it is building out the surrounding tooling — a hosted Llama API (Python/TypeScript SDKs), the PurpleLlama/Llama-Guard safety stack, and developer cookbooks — and its public…