cd/sources/pub-auto-discovered· home› sources› Pub (auto-discovered)
cat /sources/pub-auto-discovered.feed | wc -l → 1550

Pub (auto-discovered)

articles 1550 domain pub.towardsai.net → page 49/78 feed RSS
16:31
2026-07-14
pub.towardsai.net
artificial-intelligence

OpenSearch Optimizations for Production RAG

OpenSearch is a common choice for vector search in production RAG systems, but tuning vector retrieval involves trade-offs between recall and latency. The article explains how Approximate Nearest Neig…

15:31
2026-07-14
pub.towardsai.net
artificial-intelligence

How to Build a Production-Grade RAG Pipeline

A guide published on Medium explains how to build and deploy a production-grade retrieval-augmented generation (RAG) pipeline, covering hybrid search, iterative retrieval, and evaluation with test que…

15:01
2026-07-14
pub.towardsai.net
large-language-models

How I Fine-Tuned an 8B AI Model to Reason on a Free GPU

A student fine-tuned Meta's Llama 3.1 8B model for multi-step mathematical reasoning using Unsloth, LoRA, and a 'Silent Coder' approach, all within the RAM limits of a free Google Colab instance with …

14:31
2026-07-14
pub.towardsai.net
ai-agents

How to Build Fault-Tolerant Enterprise AI Agents

A large e-commerce company's AI agent processing millions of product images overnight should store progress after each completed batch in durable storage to minimize repeated work after a worker crash…

13:01
2026-07-14
pub.towardsai.net
artificial-intelligence

Multihead Attention: How One Model Reads a Sentence 144 Ways

Multihead attention in transformer models uses 12 layers each with 12 attention heads, totaling 144 heads, to read a sentence from different angles. Each head produces a length-64 update that is conca…

12:01
2026-07-14
pub.towardsai.net
artificial-intelligence

The Open Source Counter-Strike: Running Local Coding Agents

Open-source coding agents can now run entirely on local consumer hardware such as a 2020-era NVIDIA RTX 3090, eliminating usage-based pricing from AI vendors. The setup uses a quantized Gemma4 model s…

03:02
2026-07-14
pub.towardsai.net
artificial-intelligence

Cheaper AI Models Won’t Cut Your Agent Bill. Here’s Why.

Cheaper AI model pricing does not automatically reduce agent execution costs because tool schemas and repeated tool calls can consume up to 32 times more tokens than the task itself, making tool desig…

02:55
2026-07-14
pub.towardsai.net
artificial-intelligence

AI for Data Engineers

A data engineer with eight years of experience is launching a 20-part series aimed at helping fellow data engineers understand AI concepts like embeddings, RAG, and vector databases by connecting them…

← prev page 49 / 78 next →