Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

COMPANY · ENTITY #8924

FSDP

Related event timeline, sources and context from the news index.

EVENT TIMELINE

2

MODELS · 1 SOURCE · Hugging Face

Olmo launches Olmo-core 3 — open, scalable MoE training infrastructure

Olmo has released Olmo-core 3, a redesigned open mixture-of-experts (MoE) training framework intended to scale MoE models into the trillion-parameter range while preserving computational efficiency. The stack replaces an FSDP-based implementation with a DDP-centered system, combines expert/pipeline parallelism and other optimizations, reports a 2.7× throughput uplift on a 47B-parameter MoE across eight NVIDIA B300 GPUs, and shows MXFP8 delivering about 21% higher throughput than BF16 in controlled tests.

7.0

COMPANIES · 1 SOURCE · AWS Machine Learning

Run SkyRL GRPO post-training of Qwen3‑VL‑8B on Amazon SageMaker HyperPod

Amazon demonstrates running the open-source SkyRL framework on SageMaker HyperPod (EKS) to post-train a Qwen3‑VL‑8B vision‑language model with Group Relative Policy Optimization (GRPO). Using a HyperPod Ray cluster with shared FSx storage, the workflow improved maze solve rate from 43.75% to over 95% on a fixed 64‑maze evaluation set and uses colocated vLLM inference and FSDP-sharded policy training with LoRA adapters.

6.0