Thinking Machines Releases Inkling, a 1-Trillion-Parameter Multimodal Open Model
Inkling combines native image, audio, and text processing with a 1M-token context window and sparse MoE architecture for efficient multimodal reasoning.
Inkling combines native image, audio, and text processing with a 1M-token context window and sparse MoE architecture for efficient multimodal reasoning.
Cohere's first coding-specialized model combines sparse mixture-of-experts architecture with reinforcement learning for agent-based development tasks.
JetBrains' new Mixture-of-Experts model achieves 2x speedup over dense peers while activating just 2.5B parameters per token.