Follow the latest coverage, related explainers and connected technology stories.
JetBrains explains the changes it made to Junie and the inference engine to run Qwen3.6-27B locally on M5-powered MacBook devices. The experiment shows that the performance of local coding agents depends on context management and the prefill stage as much as it does on token-generation speed.
CohereLabs has launched North-Micro-Vision-Instruct, an open-weight vision-language model with 2.4 billion parameters that supports processing images at their native resolution under an Apache 2.0 license. The model targets document and chart understanding, OCR, and custom fine-tuning applications, and also includes community support for MLX-VLM and an AutoModel recipe for deployment on NVIDIA graphics processing units.