cd /news/large-language-models/looking-for-an-open-source-llm-to-re… · home topics large-language-models article
[ARTICLE · art-102558] src=discuss.huggingface.co ↗ pub= topic=large-language-models verified=true sentiment=· neutral

Looking for an Open-Source LLM to Replace Llama 3.3 70B Versatile

A developer is seeking open-source replacements for Llama 3.3 70B Versatile after Groq shut down the model, focusing on smaller parameter sizes with comparable performance for RAG, agentic workflows, and structured output. The user is considering Qwen, Mistral, DeepSeek, Gemma, and GPT-OSS, and requests recommendations for 2-5 models with details on quality, reasoning, tool calling, JSON reliability, context length, VRAM, speed, and production deployment.

read1 min views2 publishedAug 19, 2026

Hi everyone,

I was previously using Llama 3.3 70B Versatile through Groq for my application, but this model has been shut down by Groq, so I need to replace it with another model.

I am currently looking for a fully open-source/open-weight model available on Hugging Face that would be suitable for my use case.

My application involves:

RAG-based workflows

Agentic AI workflows

Prompt engineering and prompt generation

Structured/JSON output

Tool/function calling

Good instruction following

Reasoning capability

Multilingual input/output

Production use

I am particularly interested in models with smaller parameter sizes than 70B if they can provide comparable performance.

Some models I am currently considering are:

Qwen

Mistral

DeepSeek

Gemma

GPT-OSS

Could you please recommend 2–5 open-source models that would be good replacements for Llama 3.3 70B for this type of application?

It would also be helpful if you could share your experience regarding:

Model quality compared with Llama 3.3 70B

Reasoning capability

Tool/function calling

JSON/structured output reliability

Context length

VRAM/RAM requirements

Inference speed

Production deployment experience

Recommended quantization, if applicable

I am planning to test multiple models on the same set of prompts and compare their results before selecting the final model.

Thanks in advance for your recommendations!

── more in #large-language-models 4 stories · sorted by recency
── more on @llama 3.3 70b versatile 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/looking-for-an-open-…] indexed:0 read:1min 2026-08-19 ·