Local AI in .NET: ONNX Runtime, Ollama and Foundry Local
Run LLMs locally in .NET with ONNX Runtime GenAI, Ollama and Foundry Local: small language models, GPUs and NPUs, quantization and IChatClient.
11 articles about Local & On-Device AI: in-depth .NET and AI guides, senior interview questions and AI news on DotNet AI Hub.
Run LLMs locally in .NET with ONNX Runtime GenAI, Ollama and Foundry Local: small language models, GPUs and NPUs, quantization and IChatClient.
WinUI 3 and the Windows App SDK explained: packaged vs unpackaged apps, windowing, app lifecycle, Fluent controls, Windows AI APIs and deployment.
NVIDIA signed a deal to buy Hugging Face for about $12.9 billion and pledged to keep the hub open. What it means for developers who rely on open models.
Alibaba's Qwen team released Qwen3.8 in August 2026, its first Qwen-Max-class open-weight model: a 2.4T MoE with 95B active parameters, followed by a 27B model.
Dario Amodei said on July 27, 2026 that Anthropic has never sought a ban on open-weights models and backs chip controls and safety tests for all capable models.
Google DeepMind released Gemma 4 in April 2026, multimodal open models from E2B to 31B, and moved Gemma to the Apache 2.0 license for the first time.
Mistral AI released Mistral Small 4 in March 2026, an Apache 2.0 open-weight 119B MoE model with 6.5B active parameters that unifies chat, reasoning and coding.
Apple and Google agreed that Gemini will underpin the next Apple Foundation Models and a more personalized Siri. What the partnership means for developers.
OpenAI released gpt-oss-120b and gpt-oss-20b in August 2025, its first open-weight language models since GPT-2, under Apache 2.0 for reasoning and tool use.
Alibaba's Qwen team released Qwen3 in April 2025: eight Apache 2.0 open-weight models, from 0.6B to a 235B MoE, that switch between thinking and fast modes.
DeepSeek-R1 arrived in January 2025 as an MIT-licensed open-weight reasoning model that DeepSeek says matches OpenAI o1, and it shook AI markets within a week.