From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix
Paper recorded by Signals 4 on 2026-09-01 in cs.CL. Abstract reproduced from arXiv; link to the original below.
Published 2026-09-01 on arXiv · recorded by Signals 4 on 2026-09-02
Category: cs.CL · 自然语言处理 · first seen 2026-09-02
Abstract
Data-residency constraints force enterprises to self-host LLMs, but continuous adoption of newer models without decommissioning their predecessors expands the serving fleet, fragmenting a finite GPU pool. We consolidate traffic from over 200 internal applications onto a single model by closing quality gaps identified through production error analysis along three axes: instruction following, functi
Read on arXiv →
Cite this page: From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix: the #140 most recent of 186 cs.CL papers we have recorded (as of 2026-09-01). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/from-production-traffic-to-post-training-building-a-self-hosted-llm-that-covers-.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable:
papers.json
Get 4 AI signals a day by email — free.
Get 4 AI signals a day by email — free