SiliconFlow is building a "Token Factory" — an AI inference infrastructure layer that normalizes heterogeneous compute into standardized token output. The GitHub evidence reveals a…
siliconflow/glauth
forked from glauth/glauth
Forks expose upstream dependencies and research-adjacent tooling before they show up in polished launch posts. They are a quiet signal for infra, eval, agent, and model-adjacent work.
SiliconFlow leads this loaded window with 76 signals. The most repeated upstream dependency is QwenLM/Qwen3-ASR. Latest signal: Nebius - nebius/istio - forked from istio/istio.
SiliconFlow is building a "Token Factory" — an AI inference infrastructure layer that normalizes heterogeneous compute into standardized token output. The GitHub evidence reveals a…
siliconflow/glauth
forked from glauth/glauth
Snowflake is executing a deliberate convergence play: its Arctic model family — specialized for SQL, code generation, and enterprise retrieval — is being positioned not as a standalone…
Snowflake-Labs/nifi
forked from apache/nifi
Baseten is a neocloud inference platform in post-Series-F scale-up mode: nearly every job posting cites a recent $1.5B Series F led by Altimeter Capital, Conviction Partners, and Spark…
basetenlabs/go-containerregistry
forked from google/go-containerregistry
Together AI is consolidating its positioning as the AI-native cloud — an inference-first infrastructure platform that competes on raw speed and cost per token. The evidence pack shows the…
togethercomputer/GPTQModel
forked from ModelCloud/GPTQModel
Novita AI is executing a two-pronged evolution: it operates a commercial model API and agent sandbox platform for third-party frontier models, while simultaneously building deep inference…
novitalabs/tau2-bench
forked from sierra-research/tau2-bench
DeepInfra is an inference-cloud provider exploiting the open-weight model boom, not a model-building lab. Its GitHub footprint reveals a company systematically forking and maintaining the…
deepinfra/pi
forked from earendil-works/pi
DigitalOcean (GradientAI) is executing a concentrated pivot into agentic AI infrastructure as a managed cloud service, building the full stack from GPU inference to hosted agent runtimes.…
digitalocean/go-diskfs
forked from diskfs/go-diskfs
Scaleway is executing a three-pillar strategy to differentiate as Europe's sovereign AI cloud: (1) serving frontier open-weight models through its Generative APIs platform as a managed…
scaleway/winget-pkgs
forked from microsoft/winget-pkgs
CoreWeave is executing a deliberate pivot from a pure-play GPU neocloud into a vertically integrated AI platform company. The evidence shows the firm layering a proprietary…
coreweave/substrate
forked from agent-substrate/substrate
FriendliAI is an AI inference infrastructure company entering an aggressive commercialization phase, signaled by a $20M funding round, a rapid SDK iteration cadence with breaking API…
friendliai/pi-mono
forked from earendil-works/pi
Fireworks AI is a Series C ($4B valuation) generative AI infrastructure platform transitioning from inference-speed leader to full-stack AI cloud provider, with training, fine-tuning,…
fw-ai/tokenspeed
forked from lightseekorg/tokenspeed
Nebius is executing a multi-front AI cloud scaling thesis: it is simultaneously building out physical data center capacity across the US and Europe, expanding its GPU orchestration software…
nebius/istio
forked from istio/istio
Parasail is an early-stage AI infrastructure company (Series A, $32M raised, $160M valuation) building a serverless inference cloud that aggregates distributed GPU supply into an…
parasail-ai/search-grounded-ai-pipeline
forked from pc1438/search-grounded-ai-pipeline
Public AI is not a frontier model lab; it is a neocloud-adjacent public infrastructure play building an inference utility positioned as an open, democratically governed alternative to…
forpublicai/mobile-app-2
forked from cogwheel0/conduit
Replicate is a post-acquisition platform operating as an inference API aggregator, not a model builder. Following its acquisition by Cloudflare, its activity centers on platform engineering…
replicate/otel-cf-workers
forked from pydantic/otel-cf-workers
Groq is rebuilding as a pure-play AI inference cloud after a transformative non-acquisition by Nvidia that took its founding CEO, president, and key engineers. A $650M raise in June 2026…
groq/opentelemetry-collector-contrib
forked from open-telemetry/opentelemetry-collector-contrib
Wafer is a hardware-centric AI inference platform building competitive advantage through GPU kernel optimization expertise, with a distinctive multi-vendor strategy spanning NVIDIA and AMD…
wafer-ai/aiter
forked from ROCm/aiter
Hyperbolic is in a post–Series A scaling sprint, pivoting from its early Web3/decentralized microservices roots into a full-stack GPU marketplace aggregator. The evidence shows a company…
HyperbolicLabs/skypilot
forked from skypilot-org/skypilot
Cerebras is consolidating into a speed layer for frontier inference, not a model house. The through-line across the pack is that wafer-scale hardware removes the GPU memory wall and the…
Cerebras/capi-image-builder
forked from kubernetes-sigs/image-builder
Lightning AI is in the midst of a structural transformation from developer-framework shop into a vertically integrated neocloud. The merger with Voltage Park [P2, P3] has reshaped the…
Lightning-AI/skypilot
forked from skypilot-org/skypilot
Kuaishou's StreamLake is executing a dual-pronged open-source strategy: a video-native multimodal foundation model line (Keye-VL-2.0) aimed at long-video understanding and agentic…
kwaipilot/experiments
forked from SWE-bench/experiments
Blackbox AI is not a frontier model builder; it is an inference-infrastructure and agent-orchestration platform that competes on serving others' models faster, cheaper, and more securely…
No recent topic signal.
Clarifai's observable footprint is that of a platform being folded into a neocloud, not a frontier model lab: an active multi-language SDK/API release program (Python, Node.js, and gRPC…
No recent topic signal.
Cloudflare's evidence in this pack reads as a neocloud consolidating two positions at once: the security/connectivity layer for an increasingly hostile, agent-driven Internet, and the…
No recent topic signal.
Multiverse Computing is executing a deliberate pivot from its quantum-software heritage into a practical AI model-compression and deployment platform under the CompactifAI brand. The…
No recent topic signal.
Databricks in this evidence window is executing a three-phase enterprise platform consolidation: (1) an internal SAP S/4HANA transformation with an agentic AI layer that doubles as both…
No recent topic signal.
Eigen AI is a 2025-founded inference optimization company acquired by NASDAQ-listed neocloud Nebius for $643M, with the deal closing on 10 June 2026. The lab's optimization stack is being…
No recent topic signal.
GMI Cloud is an inference-optimized neocloud building a full-stack platform tightly coupled to NVIDIA's hardware roadmap. The evidence shows a company transitioning from bare-metal GPU…
No recent topic signal.
Makora is a performance-engineering organization focused on automated GPU kernel generation and inference optimization. Its public surface spans four categories: (1) an AI-driven kernel…
No recent topic signal.
SambaNova Systems is executing a decisive pivot from AI training hardware toward becoming an inference cloud provider purpose-built for agentic AI workloads. The evidence pack captures a…
No recent topic signal.