CoreWeave is executing a deliberate pivot from a pure-play GPU neocloud into a vertically integrated AI platform company. The evidence shows the firm layering a proprietary agent-development platform (W&B Weave), an autonomous research…
Neocloud intelligence
A dense desk for AI and GPU cloud providers: hiring, releases, infrastructure repos, market discussion, and the public context around capacity buildout.
What 715 open infrastructure/systems roles reveal about the GPU buildout — physical first — for two audiences: people who want to get hired in infra, and people who want to sell infra to the labs and neoclouds.
What 86 open human-data / annotation / data-quality roles reveal about the fuel layer — the clearest "sell to the labs" buy signal, since labs structurally buy data rather than build it.
Databricks in this evidence window is executing a three-phase enterprise platform consolidation: (1) an internal SAP S/4HANA transformation with an agentic AI layer that doubles as both dogfooding and product blueprint; (2) a…
- databricks/appkit
Cloudflare is running a two-sided bet that the "agentic Internet" is its next platform cycle. On one side, it is selling itself as the connectivity and security fabric for AI agents — governing MCP server sprawl and unmanaged agents via…
- cloudflare/workerd
- cloudflare/terraform-provider-cloudflare
- cloudflare/computer
DigitalOcean (GradientAI) is executing a concentrated pivot into agentic AI infrastructure as a managed cloud service, building the full stack from GPU inference to hosted agent runtimes. The evidence reveals a coordinated three-pronged…
- digitalocean/digitalocean-cloud-controller-manager
- digitalocean/doctl
- digitalocean/droplet-agent
Together AI is consolidating its positioning as the AI-native cloud — an inference-first infrastructure platform that competes on raw speed and cost per token. The evidence pack shows the company simultaneously building out in three…
- togethercomputer/together-typescript
- togethercomputer/together-py
Baseten is expanding from an inference-serving layer into three adjacent fronts at once: a capital-intensive compute-capacity operation, an agentic execution platform, and an open-source research lab. The clearest signal is hiring — the…
- basetenlabs/truss
SambaNova Systems is executing a decisive pivot from AI training hardware toward becoming an inference cloud provider purpose-built for agentic AI workloads. The evidence pack captures a company compressing its stack around three…
Lightning AI is in the midst of a structural transformation from developer-framework shop into a vertically integrated neocloud. The merger with Voltage Park [P2, P3] has reshaped the company's operational DNA: it now owns and operates…
- Lightning-AI/pytorch-lightning
- Lightning-AI/LitServe
- Lightning-AI/sdk
Nebius is executing a multi-front AI cloud scaling thesis: it is simultaneously building out physical data center capacity across the US and Europe, expanding its GPU orchestration software stack, commercializing a new agentic search…
- nebius/terraform-provider-nebius
- nebius/nebius-ps-services
- nebius/slurm-deb-packages
Snowflake is executing a deliberate convergence play: its Arctic model family — specialized for SQL, code generation, and enterprise retrieval — is being positioned not as a standalone frontier contender but as the AI inference layer…
Public AI is not a frontier model lab; it is a neocloud-adjacent public infrastructure play building an inference utility positioned as an open, democratically governed alternative to commercial AI APIs. The org's GitHub activity reveals a…
- forpublicai/chat.publicai.co
- forpublicai/chat.publicai.co
- forpublicai/chat.publicai.co
Novita AI is executing a two-pronged evolution: it operates a commercial model API and agent sandbox platform for third-party frontier models, while simultaneously building deep inference infrastructure—most visibly pegaflow, a Rust-based…
- TypeScript
- novitalabs/pegaflow
- novitalabs/pegaflow
Fireworks AI is a Series C ($4B valuation) generative AI infrastructure platform transitioning from inference-speed leader to full-stack AI cloud provider, with training, fine-tuning, serverless and dedicated inference, multi-LoRA serving,…
Scaleway is executing a three-pillar strategy to differentiate as Europe's sovereign AI cloud: (1) serving frontier open-weight models through its Generative APIs platform as a managed alternative to proprietary hyperscalers, (2) embedding…
- scaleway/ultraviolet
- scaleway/ultraviolet
- scaleway/ultraviolet
Hyperbolic is in a post–Series A scaling sprint, pivoting from its early Web3/decentralized microservices roots into a full-stack GPU marketplace aggregator. The evidence shows a company simultaneously hiring for infrastructure depth (GPU…
- HyperbolicLabs/hyperbolic-ts
- HyperbolicLabs/hyperbolic-ts
- Python
SiliconFlow is building a "Token Factory" — an AI inference infrastructure layer that normalizes heterogeneous compute into standardized token output. The GitHub evidence reveals a two-track product strategy: (1) a deep acceleration stack…
- siliconflow/aliyun-maxcompute-data-collectors
- siliconflow/aliyun-maxcompute-data-collectors
- siliconflow/aliyun-maxcompute-data-collectors
FriendliAI is an AI inference infrastructure company entering an aggressive commercialization phase, signaled by a $20M funding round, a rapid SDK iteration cadence with breaking API changes across all serving tiers, the launch of a public…
- friendliai/friendli-python
- Python
- friendliai/friendli-python
DeepInfra is an inference-cloud provider exploiting the open-weight model boom, not a model-building lab. Its GitHub footprint reveals a company systematically forking and maintaining the full inference-serving stack — from CUDA kernels to…
- deepinfra/deepinfra-node
- Python
- deepinfra/kv-local-indexer
Cerebras is consolidating into a speed layer for frontier inference, not a model house. The through-line across the pack is that wafer-scale hardware removes the GPU memory wall and the multi-chip networking tax, letting Cerebras serve the…
Groq is rebuilding as a pure-play AI inference cloud after a transformative non-acquisition by Nvidia that took its founding CEO, president, and key engineers. A $650M raise in June 2026 aims to scale GroqCloud to 200MW by 2027 and serve…
- groq/groq-python
- groq/groq-typescript
Replicate is a post-acquisition platform operating as an inference API aggregator, not a model builder. Following its acquisition by Cloudflare, its activity centers on platform engineering — evidenced by an intense cog release cadence…
- replicate/cog
Multiverse Computing is executing a deliberate pivot from its quantum-software heritage into a practical AI model-compression and deployment platform under the CompactifAI brand. The evidence depicts an organization scaling enterprise…
- Python
- Python
- CompactifAI/workshopsJan 27
Parasail is an early-stage AI infrastructure company (Series A, $32M raised, $160M valuation) building a serverless inference cloud that aggregates distributed GPU supply into an OpenAI-compatible platform for open-weight models. The…
- forked from pc1438/search-grounded-ai-pipeline
- parasail-ai/kueueJul 2forked from kubernetes-sigs/kueue
- Python
Clarifai's observable footprint is that of a platform being folded into a neocloud, not a frontier model lab: an active multi-language SDK/API release program (Python, Node.js, and gRPC clients across five languages) with essentially no…
- Clarifai/clarifai-python-grpc
- Clarifai/clarifai-php-grpc
- Clarifai/clarifai-swift-grpc
Wafer is a hardware-centric AI inference platform building competitive advantage through GPU kernel optimization expertise, with a distinctive multi-vendor strategy spanning NVIDIA and AMD accelerators. The evidence depicts a company…
- wafer-ai/wafer-docsMay 7MDX
- wafer-ai/kernel-arenaMar 10Python
- wafer-ai/aiterJan 26forked from ROCm/aiter
Makora is a performance-engineering organization focused on automated GPU kernel generation and inference optimization. Its public surface spans four categories: (1) an AI-driven kernel generation system (MakoraGenerate) that produces…
- makora-ai/makora v1.0.4Mar 30makora-ai/makora
- makora-ai/gpuq v1.5.5Feb 28makora-ai/gpuq
- makora-ai/gpuq v1.5.4Feb 28makora-ai/gpuq
Kuaishou's StreamLake is executing a dual-pronged open-source strategy: a video-native multimodal foundation model line (Keye-VL-2.0) aimed at long-video understanding and agentic capabilities, and an AI coding product suite (KAT-Coder,…
- Python
- kwaipilot/KAT-CoderSep 16HTML
Blackbox AI is not a frontier model builder; it is an inference-infrastructure and agent-orchestration platform that competes on serving others' models faster, cheaper, and more securely than anyone else. The company's public signals…
- HTML
Eigen AI is a 2025-founded inference optimization company acquired by NASDAQ-listed neocloud Nebius for $643M, with the deal closing on 10 June 2026. The lab's optimization stack is being integrated into Nebius Token Factory to deliver…
No recent non-job signal.
GMI Cloud is an inference-optimized neocloud building a full-stack platform tightly coupled to NVIDIA's hardware roadmap. The evidence shows a company transitioning from bare-metal GPU provisioning to a managed platform layer: 10 open…
No recent non-job signal.
Agent answer
handoff jsonNeocloud has 30 tracked organizations with 16,321 loaded public signals: 4,796 hiring, 670 forks, 7,378 releases or model cards, 2,135 talking, and 1,342 repos. Cloudflare (Workers AI) leads this desk with 2,985 loaded signals. Latest contextified signal: CoreWeave - 5 Misunderstandings About Enterprise AI Training Infrastructure. 30 standing agent analyses are generated for this category.
- Neocloud
tracks 30 organizations
- Neocloud
has loaded 16,321 public signals
- Neocloud
has hiring signal count 4,796
- Neocloud
has fork signal count 670
Inspect Neocloud as the category-level intelligence desk with 30 tracked organizations and 16,321 loaded public signals. Rank priority accounts across hiring (4,796), forks (670), releases (7,378), talking (2,135), and repos (1,342). Identify account momentum, source gaps, and analysis-agent dispatch work for this category.
Cloudflare (Workers AI)Cloudflare (Workers AI)2,985 signals