Spinning Up in Deep RL: Workshop review
On February 2, we held our first Spinning Up Workshop as part of our new education initiative at OpenAI.
เจาะลึกสเปกโมเดลใหม่, เปรียบเทียบ Benchmark, เทคนิคการลดต้นทุน Token และคู่มือการเชื่อมต่อ API สำหรับทีมวิศวกรไทย
On February 2, we held our first Spinning Up Workshop as part of our new education initiative at OpenAI.
We’ve written a paper arguing that long-term AI safety research needs social scientists to ensure AI alignment algorithms succeed when actual humans are involved. Properly aligning advanced AI systems with human values requires resolving many uncertainties related to the psychology of human rat...
We’ve discovered that the gradient noise scale, a simple statistical metric, predicts the parallelizability of neural network training on a wide range of tasks. Since complex tasks tend to have noisier gradients, increasingly large batch sizes are likely to become useful in the future, removing...
We’re releasing CoinRun, a training environment which provides a metric for an agent’s ability to transfer its experience to novel situations and has already helped clarify a longstanding puzzle in reinforcement learning. CoinRun strikes a desirable balance in complexity: the environment is s...
We’ve developed an energy-based model that can quickly learn to identify and generate instances of concepts, such as near, above, between, closest, and furthest, expressed as sets of 2d points. Our model learns these concepts after only five demonstrations. We also show cross-domain transfer: w...
We’ve developed Random Network Distillation (RND), a prediction-based method for encouraging reinforcement learning agents to explore their environments through curiosity, which for the first time exceeds average human performance on Montezuma’s Revenge.
We’re proposing an AI safety technique called iterated amplification that lets us specify complicated behaviors and goals that are beyond human scale, by demonstrating how to decompose a task into simpler sub-tasks, rather than by providing labeled data or a reward function. Although this idea ...
We are now accepting applications for OpenAI Fellows and Interns for 2019.
เชื่อมต่อ GPT-4.5, Claude 3.7, DeepSeek-V3 ผ่าน API เดียว พร้อมระบบ Fallback ป้องกัน Downtime และชำระเงินสะดวกผ่าน PromptPay