OngoingOnline · Tech · advanced · ai, llm, optimization
A global team-based competition to optimize SGLang AI inference kernels for multiple chips using Triton or Triton-TLE. Participants tackle over 200 performance challenges to improve LLM throughput and latency, competing for awards based on optimization breakthroughs and extreme p