cd/entity/Together· home entities Together
grep -l @together /news/*.json | wc -l → 16

Together

mentions 16 type Organization feed RSS

// recent coverage 16 mentions

18:00
2026-07-18
gist.github.com
artificial-intelligence

tutorial-kimi-k3.md

Moonshot AI's Kimi K3 model matches Opus 4.8 on intelligence benchmarks while costing ~70% less, with a 1M-token context window and open weights. The API is OpenAI-compatible, enabling easy integratio…

09:20
2026-07-16
dev.to
large-language-models

Inkling MoE + Agent Safety: Token Efficiency Meets Reliability

Inkling, a decoder-only mixture-of-experts model with 1 trillion total parameters and 40 billion active per token, launched on Together Serverless, supporting native multimodal I/O and a reasoning_eff…

00:00
2026-07-16
together.ai
artificial-intelligence

What does 99.9% uptime mean for inference?

Together, which runs inference for Cursor, Decagon, Cartesia, and Yutori, explains that 99.9% inference uptime requires surviving a full data center failure through multi-DC deployment with live traff…

22:09
2026-07-15
testingcatalog.com
artificial-intelligence

Thinking Machines debuts open-weight Inkling AI model

Thinking Machines Lab released Inkling, its first open-weights foundation model, a 975-billion-parameter Mixture-of-Experts transformer with 41 billion active parameters, available for fine-tuning via…

09:55
2026-07-04
dev.to
large-language-models

Scaling LLMs: Why Deterministic Hashing Isn't Enough

A developer built a Go library for semantic LLM caching that combines deterministic hashing with vector similarity search to reduce costs from repeated but differently worded queries. The library supp…

15:04
2026-06-28
dev.to
large-language-models

Branch Agent: Git-Style Branching for LLM Conversations

A developer built Branch Agent, a system that applies Git-style branching to LLM conversations, enabling users to fork, branch, and merge AI conversations with different models, prompts, and providers…

00:00
2026-05-08
together.ai
ai-agents

Deploy and inference any model from HuggingFace

Netflix released void-model on Hugging Face, and a developer used the Goose CLI agent with Together's dedicated containers skill to deploy the model for inference on release day with a single prompt. …

// co-occurs with top 8 entities
// topics top 6 topics