cd/entity/InstructGPT· home entities InstructGPT
grep -l @instructgpt /news/*.json | wc -l → 6

InstructGPT

mentions 6 type Organization feed RSS

// recent coverage 6 mentions

15:06
2026-07-25
promptcube3.com
large-language-models

How LLMs Defend Against Jailbreak and Prompt Injection

Prompt injection manipulates an LLM's instructions to ignore original constraints, while jailbreaking bypasses safety filters to produce restricted content. Reinforcement Learning from Human Feedback …

06:36
2026-06-19
pub.towardsai.net
large-language-models

Teaching Machines to Be Better: A Deep Dive into RLAIF and PPO

Researchers are advancing AI alignment by using Reinforcement Learning from AI Feedback (RLAIF) with Proximal Policy Optimization (PPO) to train language models, replacing expensive human annotations …

01:08
2026-06-16
dev.to
large-language-models

RLHF vs DPO vs IPO vs KTO: which alignment method should you use

A developer compares four dominant alignment methods—RLHF, DPO, IPO, and KTO—for fine-tuning large language models, detailing their mathematical formulations, data requirements, and practical tradeoff…

// co-occurs with top 8 entities
// topics top 6 topics