Frontier labfresh 3w

ByteDance (Doubao/Seed)

Signal timeline114 total
Jul 3, 2026
3wModelByteDance-Seed/PARByteDance Seed team new model release.sourcenotability 7.0/107
Jun 2, 2026
Jun 2ModelByteDance-Seed/TaskMemLow traction model release from ByteDancesourcenotability 3.0/10334
May 19, 2026
May 19ModelByteDance-Seed/SimArtNotable model release from ByteDance, but not a flagship/frontiersourcenotability 6.0/104
May 15, 2026
May 15ModelByteDance-Seed/Cola-DLMNew model from ByteDance; potentially notable.sourcenotability 7.0/1016943
Jan 15, 2026
Jan 15ModelByteDance-Seed/Stable-DiffCoder-8B-InstructNew instruction-tuned code model, modest traction.sourcenotability 5.0/10170139
Jan 15ModelByteDance-Seed/Stable-DiffCoder-8B-BaseNew model release but low tractionsourcenotability 5.0/1013320
Jan 6, 2026
Jan 6ModelByteDance-Seed/VINCIE-7BNotable model release from major companysourcenotability 7.0/1012
Dec 24, 2025
Dec 24ModelByteDance-Seed/cryofm-v2Very low HF downloadssourcenotability 2.0/10256
Dec 23, 2025
Dec 23ModelByteDance-Seed/cryofm-v1Low traction model release by ByteDance.sourcenotability 3.0/10126
Dec 2, 2025
Dec 2ModelByteDance-Seed/Adversarial-Flow-ModelsNotable research from major company, but not a flagship model.sourcenotability 6.0/1015
Dec 1, 2025
Dec 1ModelByteDance-Seed/ConfRover-interp-20M-v1.0New interpretability model from ByteDance, modest size.sourcenotability 5.0/106
Dec 1ModelByteDance-Seed/ConfRover-base-20M-v1.0Routine model release by notable company, no strong tractionsourcenotability 4.0/104
Oct 8, 2025
Oct 8ModelByteDance-Seed/AHN-Mamba2-for-Qwen-2.5-Instruct-14BLow downloads, minor variant releasesourcenotability 1.0/109115
Oct 8ModelByteDance-Seed/AHN-Mamba2-for-Qwen-2.5-Instruct-7BLow downloads, niche adaptationsourcenotability 3.0/10565
Oct 8ModelByteDance-Seed/AHN-Mamba2-for-Qwen-2.5-Instruct-3BLow downloads, niche adaptation model.sourcenotability 4.0/10737
Oct 8ModelByteDance-Seed/AHN-GDN-for-Qwen-2.5-Instruct-14BLow traction fine-tune releasesourcenotability 3.0/10536
Oct 8ModelByteDance-Seed/AHN-GDN-for-Qwen-2.5-Instruct-7BLow traction, routine fine-tunesourcenotability 3.0/10551
Oct 8ModelByteDance-Seed/AHN-GDN-for-Qwen-2.5-Instruct-3BRoutine model release with very low traction (56 downloads)sourcenotability 3.0/10681
Oct 8ModelByteDance-Seed/AHN-DN-for-Qwen-2.5-Instruct-14BLow traction, minor release from ByteDancesourcenotability 4.0/10573
Oct 8ModelByteDance-Seed/AHN-DN-for-Qwen-2.5-Instruct-7BLow traction (52 downloads), routine releasesourcenotability 2.0/10551
Oct 8ModelByteDance-Seed/AHN-DN-for-Qwen-2.5-Instruct-3BLow downloads, routine adapter releasesourcenotability 2.0/10501
Oct 6, 2025
Oct 6ModelByteDance-Seed/BFS-Prover-V2-7BModest model release, low tractionsourcenotability 5.0/104377
Sep 30, 2025
Sep 30ModelByteDance-Seed/BFS-Prover-V2-32BLow traction specialized model releasesourcenotability 3.0/106013
Aug 26, 2025
Aug 26ModelByteDance-Seed/byteff2Notable model release by ByteDancesourcenotability 7.0/105
Aug 26ModelByteDance-Seed/bamboo_mixerNew ByteDance model releasesourcenotability 6.0/1010

Top signals

  1. #1ModelsByteDance-Seed/byteff27.0
  2. #2ModelsByteDance-Seed/Cola-DLM7.0
  3. #3ReposByteDance-Seed/Depth-Anything-37.0
  4. #4ReleasesByteDance-Seed/JoltQC v0.17.0
  5. #5ModelsByteDance-Seed/PAR7.0

Agent answer

ByteDance (Doubao/Seed) has 114 loaded public signals: 0 hiring, 1 forks, 52 releases or model cards, 0 talking, and 61 repos. Latest signal: ByteDance-Seed/PAR. Data-business radar maps 4 signals to Data demand, Evals and quality, Infrastructure, Safety and policy. The standing analysis was generated with deepseek-v4-pro and 93 evidence refs.

ByteDance (Doubao/Seed)

has loaded 114 public signals

ByteDance (Doubao/Seed)

has hiring signal count 0

ByteDance (Doubao/Seed)

has fork signal count 1

ByteDance (Doubao/Seed)

has release signal count 52

Analysis — agent synthesisfull report →generated July 4, 2026

Thesis

ByteDance Seed is executing a deliberate multi-frontier strategy: shipping production models through Volcano Engine (Doubao 2.1 Pro, Seedance 2.5) while simultaneously open-sourcing a broad portfolio of research artifacts across LLM reasoning, multimodal understanding, video, biology, agents, and systems infrastructure. The evidence reveals a lab that is scaling inference infrastructure aggressively (VeOmni, Triton-distributed, ShadowKV, FlexPrefill, ByteCheckpoint), investing in agentic and VLA-world-model R&D W3W5, and releasing evaluation benchmarks that double as market-positioning instruments (EvaLearn, EdgeBench, DAComp). The hiring posture has shifted from aggressive external recruitment to internal talent development with selective elite additions W5. The absence of hiring-focused evidence in this pack is itself a signal — the lab prioritizes public releases and repos over job-post visibility.

Signal desks

Hiring

  • Seed was established in 2023 with labs across China, Singapore, and the US, spanning LLM, speech, vision, world models, AI infrastructure, and next-generation interfaces P21.
  • In 2025, ByteDance AI Lab (led by Li Hang) was merged into Seed along with its robotics team to improve coordination between models and embodied intelligence W3.
  • VLA (vision-language-action) world-model research was initiated in 2025 as a small group led by Li Hang (simulation data) and Wang Wenqian (natural data) W3.
  • The era of aggressive, high-salary external hiring is described internally as over; focus is now internal talent development with selective elite recruiting, including a former core DeepSeek researcher and a former NVIDIA research scientist W5.
  • No open job listings are cited in this evidence pack, limiting visibility into current requisition volumes or specific role themes.

Forks

  • ByteDance-Seed/triton — forked from triton-lang/triton (OpenAI's Triton compiler) on 2025-08-28; 1 star; Python/MIT. This fork is adjacent to Seed's own Triton-distributed distributed compiler project (1,457 stars) and suggests ongoing adaptation of upstream Triton for their parallel-systems compilation work E60P13.
  • No other fork activity is cited in this evidence pack.

Releases

  • Seed-OSS-36B family (Aug 2025): 36B-parameter text-generation models released under Apache 2.0. Instruct variant hit 38,057 downloads and 503 likes on Hugging Face, the highest traction release in the pack. Base and woSyn variants also released E1E4E6.
  • Seed-Thinking-v1.5 (Apr 2025): 20B-active/200B-total MoE reasoning model achieving 86.7 on AIME 2024, 55.0 on Codeforces, 77.3 on GPQA; surpasses DeepSeek R1 by 8% win rate on non-reasoning tasks. Repo stars: 812 P14E51.
  • Bagel (Apr 2025): Unified multimodal understanding/generation model, 7B-MoT variant on Hugging Face. 6,000 GitHub stars — the highest of any Seed repo P16E3.
  • Seed1.5-VL (May 2025): Vision-language model with 532M vision encoder and 20B MoE LLM; SOTA on 38/60 public benchmarks. 1,580 stars P24E37.
  • Seed-Coder (Apr 2025): 8B code LLM family (base, instruct, reasoning) using model-centric data curation. 754 stars P17E29.
  • VeOmni (Mar 2025–May 2026): Distributed training framework for any-modality models. 2,003 stars; active release cadence with versions v0.1.8 through v0.1.11 P12E17E18E23E26E30E33E34E48E59.
  • SeedVR/SeedVR2 (Jun 2025–Jan 2026): Video restoration diffusion transformer; CVPR 2025 Highlight + ICLR 2026. 1,222 stars P25E43.
  • VideoWorld series (Jan 2025): Learns world models from unlabeled video; CVPR 2025 + CVPR 2026. 790 stars P9E53.
  • Seed-Prover (Jul 2025): Formal theorem proving; solved 4/6 IMO 2025 problems during competition. 433 stars P28E39.
  • M3-Agent (Jul 2025): Agent framework with control and memorization model variants. 1,404 stars E40E7E24.
  • Stable-DiffCoder-8B (Jan 2026): Base and Instruct code models. 581 downloads, 138 likes for Instruct E2E22.
  • PAR (Jun 2026): ICML 2026 Oral — protein autoregressive modeling via multiscale structure generation P1P2E8E11.
  • Additional model releases: TaskMem, Cola-DLM, SimArt, VINCIE-7B, BFS-Prover variants, AHN-Mamba2 variants, Adversarial-Flow-Models, cryofm-v1/v2, ConfRover variants, bamboo_mixer, cudaLLM-8B, M3-Agent-Control/Memorization E5E7E12E13E16E22E24E27E28E31E32E35E41E42E44E45E46E47E49E50E52E54.

Talking

  • Seed 2.1 Pro & Turbo launch (Jun 2026): Announced at Volcano Engine FORCE conference; positioned as agent-driven enterprise models outperforming Claude Opus 4.6 on coding, agent, and multimodal benchmarks at ~80% lower TCO W1W4. Seed 2.1 Turbo is a faster, lower-cost variant for high-frequency enterprise workloads W4.
  • Seedance 2.5 video model (previewed Jun 2026 for Jul launch): Breaks the 30-second barrier for AI video generation. Seedance 2.0 upgraded to native 4K with 10-bit color. Doubao 2.1 Pro, Seedream 5.0 Pro (image), and Seed-Audio 1.0 also announced W2.
  • Four AI priorities for 2026: World models, coding, video, and monetization (Doubao). Internal hiring philosophy shifted away from aggressive external recruitment toward internal talent development W5.
  • VLA world-model thesis: ByteDance formed a small research group in 2025 to pursue the vision-language-action route for world models, merging the AI Lab into Seed for better model-embodiment coordination W3.
  • Doubao monetization: "Doubao" subscription targeting professional users across software development, data analysis, and workflow automation was announced in early 2026 W5.

Shipping

Seed ships through two distinct channels: Volcano Engine (enterprise/commercial) and open-source (Hugging Face / GitHub). The commercial side delivered Seed 2.1 Pro, Seed 2.1 Turbo, Seedance 2.5, Seedream 5.0 Pro, and Seed-Audio 1.0 at the June 2026 FORCE conference W1W2W4. On the open-source side, the lab maintains an active cadence: 36+ Hugging Face model releases since mid-2025, 20+ public GitHub repositories, and the VeOmni distributed training framework under active iterative development (v0.1.8 → v0.1.11 across Apr–May 2026) P12E17E18E23E26E30E33E48E59. Bagel (6,000 stars), Triton-distributed (1,457 stars), Seed1.5-VL (1,580 stars), M3-Agent (1,404 stars), and SeedVR (1,222 stars) are the highest-traction open-source artifacts P16P13P24E40P25.

Research themes

1. LLM reasoning & code: Seed-Thinking-v1.5 (20B/200B MoE) demonstrates strong STEM reasoning; Seed-Coder (8B) uses model-centric data curation; Seed-Prover targets formal theorem proving (IMO-level); Stable-DiffCoder explores diffusion-based code generation P14P17P28E2.

2. Multimodal understanding & generation: Bagel unifies multimodal understanding and generation in a single model; Seed1.5-VL achieves SOTA on 38/60 VLM benchmarks; SAIL explores single-transformer vision-language learning; Seed-X targets multilingual translation P16P24P15P27.

3. Video models & world models: VideoWorld learns world models from unlabeled video (CVPR 2025/2026); SeedVR/SeedVR2 delivers video restoration via diffusion transformers; Seedance 2.5 pushes AI video generation past 30 seconds P9P25W2.

4. Agents & embodied AI: Agent-R trains agents to reflect via iterative self-training; M3-Agent provides agent control and memorization models; EdgeBench benchmarks agent learning from real-world environments; Chain-of-Action (NeurIPS 2025) models robotic manipulation trajectories; UAM studies forgetting in VLA training P10E40P3P18P4.

5. Biology & science: PAR (ICML 2026 Oral) generates protein backbone structures via multi-scale autoregression; THEMol, felis, cryofm-v1/v2, and Cola-DLM extend into molecular and structural biology P1P2E20E25E45E46E5.

6. Systems & infrastructure: VeOmni for distributed any-modality training; Triton-distributed for computation-communication overlapping; ShadowKV (ICML 2025 Spotlight) and FlexPrefill (ICLR 2025 Oral) for efficient LLM inference; ByteCheckpoint (NSDI) for unified checkpointing; SDP4Bit for 4-bit communication quantization; decoupleQ for 2-bit post-training quantization; StragglerAnalysis for training efficiency diagnostics P12P13P6P7P11P8P5P19.

7. Evaluation & benchmarks: EvaLearn (NeurIPS 2025) quantifies LLM learning capability; DAComp (ICLR 2026) benchmarks data agents; EdgeBench measures agent learning from real-world environments; ByteMorph benchmarks instruction-guided image editing P22E58P3P23.

Hiring & scaling

Explicit hiring signal in this evidence pack is thin. The Seed team profile page states labs exist across China, Singapore, and the US P21. The 2026 strategy reporting indicates hiring has pivoted from "aggressive, high-salary external hiring" to internal talent development with selective elite additions — a former DeepSeek core researcher and a former NVIDIA research scientist are named as recent recruits W5. The April 2025 merger of ByteDance AI Lab into Seed, along with its robotics team, consolidated talent under one roof for model-embodiment coordination W3. No open job requisitions, headcount targets, or specific team-expansion signals are cited in this pack.

Data-business implications

  • Evals and quality: Seed releases benchmarks that double as standardized evaluation surfaces — EvaLearn (NeurIPS 2025, LLM learning capability), DAComp (ICLR 2026, data agent lifecycle), and EdgeBench (agent learning from real-world environments) all create demand for structured evaluation data, scoring infrastructure, and leaderboard tooling P22E58P3. Seed1.5-VL's claim of SOTA on 38/60 public benchmarks underscores the lab's internal eval rigor and implies substantial benchmark-curation and scoring pipeline investment P24.
  • Infrastructure: VeOmni (2,003 stars) is a production-grade distributed training framework for any-modality models with modular design, trainer-free linear scripts, and RL trainer backend — signaling demand for accelerator-agnostic training orchestration, checkpointing (ByteCheckpoint), and scaling tooling P12P11. Triton-distributed extends OpenAI Triton for parallel systems with computation-communication overlapping, pointing to deep compiler-level infrastructure work P13. ShadowKV, FlexPrefill, SDP4Bit, and decoupleQ collectively indicate heavy investment in inference optimization (KV-cache compression, sparse attention, communication quantization, 2-bit quantization) — all relevant to inference-serving platforms and GPU-infrastructure providers P6P7P8P5.
  • Data: VideoWorld's learning from unlabeled video P9 and the VLA team's dual focus on simulation vs. natural data W3 signal demand for large-scale unsupervised video data pipelines and synthetic simulation environments. Seed-Coder's model-centric data curation approach ("Let the Code Model Curate Data for Itself") suggests investment in automated data filtering and synthesis tooling P17. The DAComp benchmark specifically evaluates data agents across the "full data intelligence lifecycle" E58.
  • Product & GTM: Doubao 2.1 Pro is available through Volcano Engine with an explicit enterprise-agent positioning — coding, agent, multimodal benchmarks — at ~80% lower TCO than Claude Opus 4.6 W4W2. Doubao monetization launched in early 2026 targeting professional users for software development, data analysis, and workflow automation W5. Seed's open-source releases (Apache 2.0 and MIT licensing) serve as both community-building and platform-adoption drivers for the Volcano Engine ecosystem.
  • Safety: No cited evidence in this pack addresses safety, alignment, or red-teaming.
  • Deployment: Seed 2.1 Turbo is described as a "faster, lower cost variant built for high frequency enterprise workloads" W4, indicating a two-tier deployment strategy (Pro for capability, Turbo for throughput/cost). The ShadowKV and FlexPrefill inference optimization work directly targets high-throughput long-context deployment P6P7.

Traction highlights

  • Bagel: 6,000 GitHub stars P16
  • Depth-Anything-3: 5,692 GitHub stars E14
  • VeOmni: 2,003 GitHub stars P12
  • Seed1.5-VL: 1,580 GitHub stars P24
  • Triton-distributed: 1,457 GitHub stars P13
  • M3-Agent: 1,404 GitHub stars E40
  • SeedVR: 1,222 GitHub stars P25
  • Seed-OSS-36B-Instruct: 38,057 Hugging Face downloads, 503 likes E1
  • Seed-Thinking-v1.5: 812 GitHub stars, 86.7 AIME 2024 P14
  • VideoWorld: 790 GitHub stars P9
  • Seed-Coder: 754 GitHub stars P17
  • EvaLearn: 431 GitHub stars, NeurIPS 2025 P22
  • Seed-Prover: 433 GitHub stars, 4/6 IMO 2025 solved P28
  • Seedance 2.5: 30-second video generation barrier broken W2
  • Doubao 2.1 Pro: ~80% lower TCO vs. Claude Opus 4.6 W4

Data-business radar

cross-lab →

4 matches · 4 active lanes

ByteDance (Doubao/Seed) has a repo signal matching data demand, infrastructure, safety and policy.