Berkeley Artificial Intelligence Research (BAIR)

नवीनतम AI समाचार, मॉडल और रिलीज़ Berkeley Artificial Intelligence Research (BAIR). ['GPT-4o mini', 'GRASP', 'LLM', 'PEVA', 'ResNet', 'Transitive RL (TRL)']

शोध 🇺🇸

RL without TD Learning: A New 'Divide and Conquer' Algorithm

Researchers from BAIR (Berkeley AI) have introduced the Transitive RL (TRL) algorithm, based on a 'divide and conquer' paradigm for off-policy reinforcement learning (RL). TRL reduces the number of Bellman recursions logarithmically, avoiding the error accumulation problems typical of TD learning. The algorithm showed superior results on complex long-horizon tasks without requiring tuning of the hyperparameter n as in n-step TD.

Berkeley Artificial Intelligence Research (BAIR)Berkeley Artificial Intelligence Research (BAIR)
BAIR (Berkeley AI)27.07 · 17:04
शोध 🇺🇸

GRASP: वर्ल्ड मॉडल्स में लंबी अवधि की योजना बनाने के लिए ग्रेडिएंट-आधारित योजनाकार

बर्कले एआई (BAIR) के शोधकर्ताओं ने GRASP पेश किया, जो डायनेमिक्स (वर्ल्ड मॉडल) सीखने के लिए एक ग्रेडिएंट-आधारित योजनाकार है, जो लंबी अवधि की योजना को व्यावहारिक बनाता है। यह ट्रैजेक्टरी को लेटेंट स्टेट्स में उठाकर समानांतर ऑप्टिमाइजेशन करता है, एक्सप्लोरेशन के लिए स्टोकैस्टिसिटी जोड़ता है, और उच्च-आयामी विजुअल मॉडलों के माध्यम से नाजुक ग्रेडिएंट से बचने के लिए ग्रेडिएंट को रूपांतरित करता है।

Berkeley Artificial Intelligence Research (BAIR)Berkeley Artificial Intelligence Research (BAIR)
BAIR (Berkeley AI)27.07 · 15:05
ताज़ा समाचार