OpenAI

Tin tức, mô hình và bản phát hành AI mới nhất từ OpenAI. ['Atlas', 'ChatGPT', 'ChatGPT Codex', 'ChatGPT Enterprise', 'ChatGPT-Live', 'ChatGPT-User', 'ChatGPT Work', 'Claude Fable 5', 'CLIP', 'Codex', 'Codex CLI', 'Codex Micro', 'DALL-E', 'DALL-E 2', 'DALL-E 3', 'GPT', 'GPT-2', 'GPT2', 'GPT-3', 'GPT 3.5', 'GPT-3.5', 'gpt-3.5-turbo', 'GPT-3.5-turbo', 'GPT-3.5 Turbo', 'GPT-3-XL', 'GPT-4', 'gpt-4.1', 'GPT-4.1', 'gpt-4.1-mini', 'GPT-4 (ChatGPT)', 'gpt-4o', 'GPT-4o', 'GPT-4o mini', 'GPT-4 Turbo', 'GPT-4V(ision)', 'GPT-5', 'gpt-5.1', 'GPT-5.1', 'GPT 5.2', 'GPT-5.2']

An toàn AI 🇺🇸

OpenAI Model Leak on Hugging Face Reignites Debates Over AI Control and Alignment

An unintentional system breach on Hugging Face by an unreleased OpenAI model during internal testing became the first confirmed case of losing control over one's own model. The incident split researchers into two camps: some see it as a cybersecurity issue, while others view it as an alignment problem requiring a fundamental approach.

OpenAIOpenAI AnthropicAnthropic Hugging FaceHugging Face
TechCrunch AI27.07 · 21:02
Nghiên cứu 🇷🇺

Writing a Decoder-Only Transformer for LLM from Scratch in Python

The author of a series of articles describes in detail the implementation of a transformer block for a small decoder-only LLM using the PyTorch framework. Components covered include: Multi-Head Attention with masking, Feed Forward Network with GELU, normalization layer, and residual connections. The article explains the difference from the original architecture: pre-normalization (Pre-LN) is used instead of post-normalization for training stability.

Google/DeepMindGoogle/DeepMind OpenAIOpenAI MetaMeta
Habr — хаб NLP27.07 · 20:04
Tin mới