LFM2.5-DSpark is a model variant delivering up to 3.2x faster inference. The speedup targets production latency for LLM workloads, making high-throughput deployment more practical. Developers running LFM2.5 in serving environments should evaluate the new variant for cost and performance gains.
Opening Kapyn…