semantic-scholar
Semantic Scholar1 events on 2025-11-19role: thematic30d fetch windowcadence: on demandaccess: keyless
Extracts: paper title, abstract snippet · metadata or bounded excerpt
Retention: D1 event history with canonical source link and deduplication metadata.
Access basis: Public endpoint; no project credential required.
Last ingested: 2026-08-12 · run status: on demand
Archive source — full history has value. Use pagination to browse older records.
Query: mixture of experts inference serving Authors: Kexin Chu, Dawei Xiang, Zixu Shen, Yiwei Yang, Zechen Liu Citations: 3 Mixture-of-Experts (MoE) has become a practical architecture for scaling LLM capacity while keeping per-token compute modest, but deploying MoE models on a