SiliconFlow is building a "Token Factory" — an AI inference infrastructure layer that normalizes heterogeneous compute into standardized token output. The GitHub evidence reveals a…
siliconflow/node_exporter
forked from prometheus/node_exporter
Forks expose upstream dependencies and research-adjacent tooling before they show up in polished launch posts. They are a quiet signal for infra, eval, agent, and model-adjacent work.
SiliconFlow leads this loaded window with 60 signals. The most repeated upstream dependency is BerriAI/litellm. Latest signal: SiliconFlow - siliconflow/node_exporter - forked from prometheus/node_exporter.
SiliconFlow is building a "Token Factory" — an AI inference infrastructure layer that normalizes heterogeneous compute into standardized token output. The GitHub evidence reveals a…
siliconflow/node_exporter
forked from prometheus/node_exporter
Baseten is in the midst of a breakout scaling phase fueled by a $1.5B Series F. The company is tripling headcount while simultaneously shipping at high velocity across three vectors:…
basetenlabs/go-containerregistry
forked from google/go-containerregistry
Snowflake is executing a deliberate convergence play: its Arctic model family — specialized for SQL, code generation, and enterprise retrieval — is being positioned not as a standalone…
Snowflake-Labs/homebrew-core
forked from Homebrew/homebrew-core
Together AI is consolidating its positioning as the AI-native cloud — an inference-first infrastructure platform that competes on raw speed and cost per token. The evidence pack shows the…
togethercomputer/metabase
forked from metabase/metabase
Novita AI is executing a two-pronged evolution: it operates a commercial model API and agent sandbox platform for third-party frontier models, while simultaneously building deep inference…
novitalabs/glm-simple-evals
forked from zai-org/glm-simple-evals
DeepInfra is an inference-cloud provider exploiting the open-weight model boom, not a model-building lab. Its GitHub footprint reveals a company systematically forking and maintaining the…
deepinfra/pi
forked from earendil-works/pi
DigitalOcean (GradientAI) is executing a concentrated pivot into agentic AI infrastructure as a managed cloud service, building the full stack from GPU inference to hosted agent runtimes.…
digitalocean/go-diskfs
forked from diskfs/go-diskfs
Scaleway is executing a three-pillar strategy to differentiate as Europe's sovereign AI cloud: (1) serving frontier open-weight models through its Generative APIs platform as a managed…
scaleway/winget-pkgs
forked from microsoft/winget-pkgs
CoreWeave is transitioning from a GPU-rental neocloud into a full-stack AI cloud platform with serious public-company scale and discipline. The evidence points to three reinforcing vectors:…
coreweave/substrate
forked from agent-substrate/substrate
FriendliAI is an AI inference infrastructure company entering an aggressive commercialization phase, signaled by a $20M funding round, a rapid SDK iteration cadence with breaking API…
friendliai/pi-mono
forked from earendil-works/pi
Groq is rebuilding as a pure-play AI inference cloud after a transformative non-acquisition by Nvidia that took its founding CEO, president, and key engineers. A $650M raise in June 2026…
groq/opentelemetry-collector-contrib
forked from open-telemetry/opentelemetry-collector-contrib
Parasail is an early-stage AI infrastructure company (Series A, $32M raised, $160M valuation) building a serverless inference cloud that aggregates distributed GPU supply into an…
parasail-ai/search-grounded-ai-pipeline
forked from pc1438/search-grounded-ai-pipeline
Fireworks AI is a Series C ($4B valuation) generative AI infrastructure platform transitioning from inference-speed leader to full-stack AI cloud provider, with training, fine-tuning,…
fw-ai/tokenspeed
forked from lightseekorg/tokenspeed
Replicate is a post-acquisition platform operating as an inference API aggregator, not a model builder. Following its acquisition by Cloudflare, its activity centers on platform engineering…
replicate/otel-cf-workers
forked from pydantic/otel-cf-workers
Nebius is executing a multi-front AI cloud scaling thesis: it is simultaneously building out physical data center capacity across the US and Europe, expanding its GPU orchestration software…
nebius/cluster-api-ipam-provider-in-cluster
forked from kubernetes-sigs/cluster-api-ipam-provider-in-cluster
Public AI is not a frontier model lab; it is a neocloud-adjacent public infrastructure play building an inference utility positioned as an open, democratically governed alternative to…
forpublicai/mobile-app-2
forked from cogwheel0/conduit
Wafer is a hardware-centric AI inference platform building competitive advantage through GPU kernel optimization expertise, with a distinctive multi-vendor strategy spanning NVIDIA and AMD…
wafer-ai/aiter
forked from ROCm/aiter
Hyperbolic is in a post–Series A scaling sprint, pivoting from its early Web3/decentralized microservices roots into a full-stack GPU marketplace aggregator. The evidence shows a company…
HyperbolicLabs/skypilot
forked from skypilot-org/skypilot
Cerebras is executing a hard pivot from wafer-scale training hardware specialist to full-stack inference cloud provider. The evidence shows a company that IPO'd in May 2026, closed an $850M…
Cerebras/capi-image-builder
forked from kubernetes-sigs/image-builder
Lightning AI is in the midst of a structural transformation from developer-framework shop into a vertically integrated neocloud. The merger with Voltage Park [P2, P3] has reshaped the…
Lightning-AI/skypilot
forked from skypilot-org/skypilot
Kuaishou's StreamLake is executing a dual-pronged open-source strategy: a video-native multimodal foundation model line (Keye-VL-2.0) aimed at long-video understanding and agentic…
kwaipilot/experiments
forked from SWE-bench/experiments
Blackbox AI is not a frontier model builder; it is an inference-infrastructure and agent-orchestration platform that competes on serving others' models faster, cheaper, and more securely…
No recent topic signal.
Clarifai is in wind-down mode as a standalone entity. The company's platform and inference IP have been acquired by Nebius (NASDAQ: NBIS), with all Clarifai services ceasing on July 17,…
No recent topic signal.
Cloudflare is executing a multi-vector pivot from content-delivery infrastructure toward AI-native compute and agentic intermediation. The evidence shows three reinforcing moves: (1)…
No recent topic signal.
CompactifAI (Multiverse Computing) is transitioning from a quantum-software R&D shop into a commercial AI infrastructure company anchored by model compression. Its proprietary CompactifAI…
No recent topic signal.
Databricks is executing a three-front offensive: it is monetising enterprise data gravity through Lakebase (a serverless Postgres for operational/AI workloads), hardening an agent platform…
No recent topic signal.
Eigen AI is a 2025-founded inference optimization company acquired by NASDAQ-listed neocloud Nebius for $643M, with the deal closing on 10 June 2026. The lab's optimization stack is being…
No recent topic signal.
GMI Cloud is an inference-optimized neocloud building a full-stack platform tightly coupled to NVIDIA's hardware roadmap. The evidence shows a company transitioning from bare-metal GPU…
No recent topic signal.
Makora is a performance-engineering organization focused on automated GPU kernel generation and inference optimization. Its public surface spans four categories: (1) an AI-driven kernel…
No recent topic signal.
SambaNova Systems is executing a decisive pivot from AI training hardware toward becoming an inference cloud provider purpose-built for agentic AI workloads. The evidence pack captures a…
No recent topic signal.