跳到主要内容
返回新闻室

Gpqa

2 条新闻

来自世界各地可信来源的关于 Gpqa 的最新新闻和动态。

OpenAI Launches o1 Pro, a More Powerful Reasoning Model for Complex Tasks

OpenAI Launches o1 Pro, a More Powerful Reasoning Model for Complex Tasks

OpenAI has introduced o1 Pro, a new reasoning model designed to deliver stronger performance on demanding tasks in science, coding, and analysis. According to the company, the model outperforms its predecessor, o1, on key benchmarks including AIME and GPQA, with a reported 94% score on AIME math problems. The release underscores OpenAI’s push to improve advanced reasoning capabilities as competition intensifies…

科技 123 次浏览 151d前
0
OpenAI Unveils o1-Pro, Its Strongest Reasoning Model to Date

OpenAI Unveils o1-Pro, Its Strongest Reasoning Model to Date

OpenAI has introduced o1-pro, a new reasoning model the company says delivers its best performance yet on complex problem-solving tasks. In benchmark testing, o1-pro outperformed the standard o1 model on challenging evaluations including AIME math, where it scored 93% versus 79%, and GPQA science, where it reached 84% compared with 74% for o1. The results position o1-pro as a more capable option for users who need…

科技 104 次浏览 152d前
0
OpenAI Launches o1 Reasoning Model, Touting Major Gains on Hard Problems

OpenAI Launches o1 Reasoning Model, Touting Major Gains on Hard Problems

OpenAI has unveiled o1, a new family of models built to tackle complex reasoning tasks in math, coding, and science. The company says the system is designed to think through problems step by step, using reinforcement learning to improve its ability to work through difficult questions before producing an answer. According to OpenAI, o1 has delivered strong benchmark results, including an 83% score on AIME 2024 and…

科技 69 次浏览 155d前
0
OpenAI Launches o1 Reasoning Model, Marking Major Leap in Complex Problem-Solving

OpenAI Launches o1 Reasoning Model, Marking Major Leap in Complex Problem-Solving

OpenAI has unveiled o1, a new family of models designed to tackle complex reasoning tasks in mathematics, coding, and science. The company says the system represents a significant step forward in AI performance, with strong results on demanding benchmarks including AIME and GPQA, where it outperformed earlier models by wide margins. Unlike traditional models that rely primarily on rapid pattern matching, o1 uses a…

科技 36 次浏览 155d前
0

快速导航