OpenAI

최신 AI 뉴스, 모델 및 출시 정보 OpenAI. ['Atlas', 'ChatGPT', 'ChatGPT Enterprise', 'ChatGPT-Live', 'ChatGPT-User', 'ChatGPT Work', 'Claude Fable 5', 'CLIP', 'Codex', 'Codex CLI', 'Codex Micro', 'DALL-E', 'DALL-E 2', 'DALL-E 3', 'GPT', 'GPT-2', 'GPT2', 'GPT-3', 'GPT 3.5', 'GPT-3.5', 'gpt-3.5-turbo', 'GPT-3.5-turbo', 'GPT-3.5 Turbo', 'GPT-3-XL', 'GPT-4', 'gpt-4.1', 'GPT-4.1', 'gpt-4.1-mini', 'GPT-4 (ChatGPT)', 'gpt-4o', 'GPT-4o', 'GPT-4o mini', 'GPT-4 Turbo', 'GPT-4V(ision)', 'GPT-5', 'gpt-5.1', 'GPT-5.1', 'GPT 5.2', 'GPT-5.2', 'GPT-5.2-codex']

모델 🇷🇺

Kimi K3 작동 원리: 2.8조 파라미터, 선형 어텐션, 백만 토큰 에이전트

Moonshot AI가 Kimi K3 모델의 가중치를 공개했습니다. 이 모델은 Mixture of Experts 아키텍처를 기반으로 한 멀티모달 모델로, 총 2.8조 개의 파라미터 중 토큰당 1040억 개가 활성화됩니다. 특정 프로그래밍 및 도구 사용 벤치마크에서 경쟁사를 능가하지만, 평균적으로는 일부 경쟁사에 뒤처집니다.

Moonshot AIMoonshot AI AnthropicAnthropic OpenAIOpenAI
Habr — хаб ИИ27.07 · 23:03
AI 안전 🇺🇸

AI Import 461: 'Alignment Not on Right Track'; FrontierCode; and Synthetic Science Interns

Researchers from UK AISI and Timaeus founded the non-profit organization Sequent to develop alignment methods ensuring the safety of superintelligent systems. Cognition released the FrontierCode benchmark for evaluating code quality, where the best model, Claude Opus 4.8, scored only 13.4% on the hardest level. Also presented are the ChinaHeritaQA dataset for evaluating VLM cultural knowledge about Chinese UNESCO sites and the Xiaomi MiMo-V2.5-Pro-UltraSpeed model with a generation speed of 1000 tokens per second.

CognitionCognition Alibaba/QwenAlibaba/Qwen AnthropicAnthropic OpenAIOpenAI
Import AI27.07 · 21:04
새로운 뉴스