Congrats to the @NexEcosystem on launching NexRT!
SGLang runs prefill and NexRT takes over decode for low-latency single-request inference on Nex-N2.5-Pro.
- The two engines share routed MoE weights
- DFlash builds on target feature capture in SGLang v0.5.15
Try NexRT with SGLang: https://github.com/nex-agi/NexRT
