Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

COMPANY · ENTITY #7966

H100

Related event timeline, sources and context from the news index.

EVENT TIMELINE

3

COMPANIES · 1 SOURCE · NVIDIA Developer

NVIDIA cuOpt adds mPDLP multi‑GPU LP solver that scales to 100M+ variables

NVIDIA released mPDLP, a Multi‑GPU Primal‑Dual hybrid gradient solver in cuOpt that shards large linear programs across NVLink‑connected GPUs. mPDLP cuts per‑GPU peak memory (up to 6× vs single‑GPU PDLP), supports problems up to 2.1B nonzeros, and in DGX B200 benchmarks shows PDLP‑step speedups up to 11.4× and typical overall gains of 1.2×–2.5× vs prior D‑PDLP; partner cases reported ~3.3× (Kinaxis, 135M variables) and >5× (PSR, 185M variables). A tutorial and source code are available on GitHub.

7.0

MODELS · 1 SOURCE · InfoQ AI, ML & Data Engineering

GKE Pod snapshots cut model startup latency up to 89%; GA on 1.35.3-gke.1234000+

Google published benchmarks for GKE Pod snapshots (general availability in May on clusters running 1.35.3-gke.1234000 or later) that checkpoint and restore full pod state including CPU/GPU memory, reporting startup latency reductions up to 89% (a 70B-parameter model in 37s, an 8B model in 15s). The feature relies on gVisor and GKE Sandbox (Autopilot has it by default), stores snapshot data in Cloud Storage, and is configured via PodSnapshotStorageConfig and PodSnapshotPolicy; compatibility rules and application rehydration limitations mean operational work shifts to snapshot lifecycle management.

7.0

MODELS · 1 SOURCE · Hugging Face

Liquid AI publishes DSpark drafter for LFM2.5-VL-3B, accelerating VLM decoding up to 3.13×

Liquid AI released a DSpark draft model for its vision‑language model LFM2.5‑VL‑3B that adds speculative decoding to speed up token decoding (up to 3.13× on-device, 2.66× on H100) while increasing model size by ~280M parameters (≈8.9%). The drafter ships with day‑one integrations for llama.cpp, MLX‑VLM and SGLang and is available on Hugging Face in Safetensors and GGUF formats; measured end‑to‑end gains ranged up to 2.62× on edge and 2.27× on GPU across MMSpec tasks.

7.0