cd/entity/TinyLlama· home entities TinyLlama
grep -l @tinyllama /news/*.json | wc -l → 18

TinyLlama

mentions 18 type Organization feed RSS

// recent coverage 18 mentions

10:43
2026-08-19
srirupa19.github.io
artificial-intelligence

Benchmarking Small Models for DigiKam's Natural Language Search

Google Summer of Code 2026 contributor benchmarking for digiKam's natural language search found that Qwen2.5-1.5B outperformed TinyLlama and a third model on real queries, with Qwen scoring 66% accura…

00:36
2026-08-19
dev.to
developer-tools

Build a Tiny Native Coding Agent in Under 100 Lines

A developer outlines how to build a lightweight, native coding agent in under 100 lines of code, using small open-source models like CodeGemma 2B or TinyLlama 1.1B and quantization for efficiency. The…

23:28
2026-08-11
runtimewire.com
artificial-intelligence

Cua ships a Metal shim that accelerates LLMs inside macOS VMs

Cua released an open-source Metal capability layer on August 11th that accelerates llama.cpp inference inside Apple Silicon macOS virtual machines, reporting prompt-processing gains up to 11.08x and t…

17:09
2026-08-11
sourcefeed.dev
artificial-intelligence

macOS VMs Have Been Sandbagging Your GPU

Cua's Metal capability shim recovers up to 16.36× llama.cpp throughput in macOS virtual machines by unmasking GPU capabilities that Apple's paravirtualized graphics device underreports. On an M1 Ultra…

14:00
2026-07-28
kdnuggets.com
large-language-models

An Introductory Guide to Practical Constraint Decoding

Practical constraint decoding, also known as structured generation, forces large language models to generate outputs that strictly follow a specified data schema, grammar, or regex by masking logits a…

12:48
2026-07-23
promptcube3.com
large-language-models

Local LLM Setup: Intel MacBook Air 2017

A user with a 2017 Intel MacBook Air seeks advice on running local LLMs, citing compatibility issues with Ollama due to macOS 12 restrictions. Potential solutions include LM Studio, llama.cpp, and GPT…

06:59
2026-07-04
dev.to
developer-tools

OrinIDE v1.0.8 is here and it's a whole vibe upgrade 🚀

OrinIDE v1.0.8, an AI-powered code editor that runs entirely in the browser without cloud accounts or subscriptions, now supports offline AI models via Ollama and introduces a 4-agent workflow for pla…

04:12
2026-07-04
dev.to
artificial-intelligence

KENSAT: A Home-Built CubeSat Running a TinyLlama LLM in Orbit

Ken Chan built KENSAT, a 2U CubeSat running a quantized TinyLlama large language model on an NVIDIA Jetson Orin Nano, scheduled to launch this fall. The satellite performs AI inference in orbit, trans…

13:01
2026-06-24
gist.github.com
large-language-models

NVIDIA GenAI LLM Certification Lab

NVIDIA has released a GenAI LLM Certification Lab that guides developers through building a production-ready fine-tuning and optimization pipeline. The lab covers data preparation, LoRA fine-tuning wi…

07:14
2026-05-20
dev.to
large-language-models

I Thought Fine-Tuning LLMs Needed Expensive GPUs. I Was Wrong.

The author successfully fine-tuned a 1.1 billion parameter TinyLlama model using QLoRA on consumer hardware, training only 0.2% of the model's parameters via low-rank adapter matrices. The project inv…

// co-occurs with top 8 entities
// topics top 6 topics