T
Tora AI Newsroom
⚡ ข่าวสารและบทวิเคราะห์เทคนิค

เกาะติดโมเดลและเทคโนโลยี AI สำหรับนักพัฒนาซอฟต์แวร์

เจาะลึกสเปกโมเดลใหม่, เปรียบเทียบ Benchmark, เทคนิคการลดต้นทุน Token และคู่มือการเชื่อมต่อ API สำหรับทีมวิศวกรไทย

บทความล่าสุด 1075 บทความ

Tech

Learning to model other minds

We’re releasing an algorithm which accounts for the fact that other agents are learning too, and discovers self-interested yet collaborative strategies like tit-for-tat in the iterated prisoner’s dilemma. This algorithm, Learning with Opponent-Learning Awareness (LOLA), is a small step toward...

Tora Technical Editorial อ่านต่อ
Tech

More on Dota 2

Our Dota 2 result shows that self-play can catapult the performance of machine learning systems from far below human level to superhuman, given sufficient compute. In the span of a month, our system went from barely matching a high-ranked player to beating the top pros and has continued to improv...

Tora Technical Editorial อ่านต่อ
Tech

Better exploration with parameter noise

We’ve found that adding adaptive noise to the parameters of reinforcement learning algorithms frequently boosts performance. This exploration method is simple to implement and very rarely decreases performance, so it’s worth trying on any problem.

Tora Technical Editorial อ่านต่อ
OPENAI

Proximal Policy Optimization

We’re releasing a new class of reinforcement learning algorithms, Proximal Policy Optimization (PPO), which perform comparably or better than state-of-the-art approaches while being much simpler to implement and tune. PPO has become the default reinforcement learning algorithm at OpenAI because...

Tora Technical Editorial อ่านต่อ
Tech

Robust adversarial inputs

We’ve created images that reliably fool neural network classifiers when viewed from varied scales and perspectives. This challenges a claim from last week that self-driving cars would be hard to trick maliciously since they capture images from multiple scales, angles, perspectives, and the like.

Tora Technical Editorial อ่านต่อ
Tech

Faster physics in Python

We’re open-sourcing a high-performance Python library for robotic simulation using the MuJoCo engine, developed over our past year of robotics research.

Tora Technical Editorial อ่านต่อ
Tech

Learning to cooperate, compete, and communicate

Multiagent environments where agents compete for resources are stepping stones on the path to AGI. Multiagent environments have two useful properties: first, there is a natural curriculum—the difficulty of the environment is determined by the skill of your competitors (and if you’re competing...

Tora Technical Editorial อ่านต่อ
Tech

Robots that learn

We’ve created a robotics system, trained entirely in simulation and deployed on a physical robot, which can learn a new task after seeing it done once.

Tora Technical Editorial อ่านต่อ

พร้อมต่อยอดโมเดล AI ในระบบของคุณแล้วหรือยัง?

เชื่อมต่อ GPT-4.5, Claude 3.7, DeepSeek-V3 ผ่าน API เดียว พร้อมระบบ Fallback ป้องกัน Downtime และชำระเงินสะดวกผ่าน PromptPay