AWS Machine LearningLabs
Scaling MoE reinforcement learning on Amazon EKS with EFA and DeepEP with 40% more throughput

When you post-train a Mixture-of-Experts (MoE) model with Reinforcement Learning from Human Feedback (RLHF) or Group Relative Policy Optimization (GRPO) at scale, three simultaneous challenges…
Read at AWS Machine Learning ↗Related

LabsEvaluate skill-equipped agents with Strands Evals and Amazon Bedrock AgentCore AWS Machine Learning

LabsBest practices guide for customizing Gemini models via Reinforcement Learning (RL) Google Cloud AI

GitHubFrom Agent Authorization to AI Production Evaluation: QCon AI New York 2026 InfoQ AI

SiliconEfficient MoE Training for Biological Foundation Models NVIDIA Technical Blog

VoicesChroma and Agentic Retrieval Software Engineering Daily
