Skip to content
ArxDecoded

ArxDecoded

  • Home
  • AI & Emerging
  • Dev & Open Source
  • Cloud & Security
  • Hardware & Chips
  • Mobile & Web

FP8

PaddlePaddle 3.1.0: Auto-Parallel, FP8, and CUDA-Like Hardware Support

August 26, 2026

PaddlePaddle 3.1.0 introduces a refined auto-parallel architecture, FP8 low-precision training for 10-20% speedups, and a mechanism to reuse CUDA kernels for heterogeneous hardware.

Categories AI & Machine Learning, Software & Open Source

PaddlePaddle 3.1.1: Enhanced Stability and FP8 Performance for Large Language Models

August 26, 2026

Baidu’s PaddlePaddle framework updates to v3.1.1, focusing on robust large model training, FP8 optimization, and new API features.

Categories AI & Machine Learning, Software & Open Source

PaddlePaddle 3.2.0: Optimizing Large Model Training and Hardware Adaptation

August 26, 2026

PaddlePaddle 3.2.0 emphasizes performance for large language models through FlashMask V3, FP8 training, and enhanced fault tolerance. It also expands hardware support via a Custom Device API and improves compatibility with industry standards like Safetensors and PyTorch-style APIs.

Categories AI & Machine Learning, Software & Open Source
  • About
  • Sources
  • Contact
  • Disclaimer
  • Privacy Policy
  • Sitemap
© 2026 ArxDecoded · Clear technology, grounded in evidence.