Liquid AI releases LFM2.5-Encoders: sub-billion-parameter models for CPU-based long-context NLP
Two new encoder models from Liquid AI match larger baselines on GLUE and SuperGLUE while maintaining 8,192-token context and 3.7× CPU speed advantage over ModernBERT-base.
GLM-5.2 Targets Long-Horizon Engineering Tasks With 1M-Token Context
Zhipu AI's GLM-5.2 delivers sustained long-context performance for multi-hour coding projects, outpacing open-source competitors on software engineering benchmarks.
SubQ Claims 12-Million-Token Context at Sub-Quadratic Cost
A new architecture called SubQ targets 12 million token context windows while sidestepping the quadratic compute scaling that limits standard transformers.