Categories

AI Agents

How to use DeepSeek Harness with a local model

DeepSeek Harness (dsh) is the open-source agent harness DeepSeek released in August 2026. It works with a local model if you declare an openai-completions provider in settings.yaml that points at llama-server, with a placeholder key and the real context size. With Qwen3.5-4B on CPU it left the tests green in 4 of 5 attempts, but slowly.

Technology

Next-generation NPUs: the hardware moving AI in 2026

NPUs stopped being an accessory and became the component that defines real performance in laptops, phones, and small servers. A practical look at the hardware that rules 2026, which workloads pay off, and where the traditional GPU still wins.

Artificial Intelligence

Phi-3 on the edge: Microsoft’s SLM in 2025

Phi-3 is Microsoft Research's family of small language models aimed at the edge, and it competes directly with Llama 3.2, Gemma 2 and Qwen 2.5. Phi-3-mini holds 3.8B parameters and, quantized to 4 bits, fits in about 2 GB, running on a CPU with a neural accelerator, an integrated GPU or an NPU.