BANANDRE
NO ONE CARES ABOUT CODE

Navigation

HomeCategories

Categories

Artificial Intelligence(619)
Software Architecture(314)
Software Development(293)
Data Engineering(174)
Engineering Management(88)
Enterprise Architecture(73)
Product Management(30)

Tagged with

#Mixture of Experts

3 articles found

K-EXAONE 2.0: LG’s 750B Sovereign AI Gamble That’s Either Genius or Overkill
K-EXAONE 2.0
Featured

K-EXAONE 2.0: LG’s 750B Sovereign AI Gamble That’s Either Genius or Overkill

LG AI Research drops a 750B-parameter MoE behemoth under Apache 2.0. It crushes long-context benchmarks but pits 37B active params against models 10x smaller. Is this Korea’s sovereign AI win or a compute flex with diminishing returns?

#K-EXAONE 2.0#LG AI Research#Mixture of Experts...
Read More
The 753B Model You Can Actually Own: GLM-5.2 Distillation Will Upend Local AI
distillation

The 753B Model You Can Actually Own: GLM-5.2 Distillation Will Upend Local AI

GLM-5.2 is the third-best model overall, but its MIT license means the real magic, distillation into small, local models, hasn’t even started yet.

#distillation#Frontier Models#GLM-5.2...
Read More
1000 Tokens Per Second on a 1T Model? Xiaomi Just Broke Physics (or At Least the Latency Barrier)
distributed systems

1000 Tokens Per Second on a 1T Model? Xiaomi Just Broke Physics (or At Least the Latency Barrier)

Xiaomi’s MiMo v2.5 hits 1000 TPS on a trillion-parameter model using commodity GPUs. Here’s the deep dive on the FP4 quantization, DFlash speculative decoding, and TileRT systems alchemy that made it possible.

#distributed systems#Inference Optimization#Mixture of Experts...
Read More
BANANDRE
NO ONE CARES ABOUT CODE

Connect

2026 BANANDRE
Privacy PolicyTermsImpressum
Built with 🍌