Skip to main content

Để giao dịch tại Mỹ, hãy truy cập polymarket.us

icon for Next Grok Model (4.6+): Humanity’s Last Exam Debut?

Next Grok Model (4.6+): Humanity’s Last Exam Debut?

icon for Next Grok Model (4.6+): Humanity’s Last Exam Debut?

Next Grok Model (4.6+): Humanity’s Last Exam Debut?

$29,129 KL.

Dec 31, 2026
Polymarket

$29,129 KL.

Polymarket

35%+

$4,090 KL.

Yes

40%+

$9,105 KL.

No

45%+

$9,706 KL.

No

50%+

$3,907 KL.

No

55%+

$2,321 KL.

No

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".**Grok 4.6, xAI’s 1.5-trillion-parameter successor to Grok 4.5, shipped in early August 2026 with post-training upgrades targeting coding, agentic workflows, and engineering tasks.** Prior Grok 4 variants posted competitive HLE results, including up to 50% on multi-agent “Heavy” configurations, though public leaderboards remain led by Anthropic’s Claude Fable 5.1 and Opus 5 at 64.5–65%. The market reflects trader focus on whether the latest Grok release will promptly appear on the official HLE leaderboard at frontier-comparable accuracy, given xAI’s pattern of benchmark emphasis and the benchmark’s rapid score climb from single-digit percentages in early 2025 to the mid-40s–60s% range today. Key near-term catalysts include any verified HLE submission, a potential Grok 4.7 follow-up, and ongoing competition in reasoning and expert-knowledge evaluations.

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No".

If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results.

A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify.

The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere.

If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered.

A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market.

The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Khối lượng
$29,129
Ngày kết thúc
Dec 31, 2026
Thị trường mở
Aug 10, 2026, 6:25 PM ET

Nguồn giải quyết

https://agi.safe.ai/

Người giải quyết

0x65070BE91...

Kết quả đề xuất: Yes

Không tranh chấp

Kết quả cuối cùng: Yes

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".**Grok 4.6, xAI’s 1.5-trillion-parameter successor to Grok 4.5, shipped in early August 2026 with post-training upgrades targeting coding, agentic workflows, and engineering tasks.** Prior Grok 4 variants posted competitive HLE results, including up to 50% on multi-agent “Heavy” configurations, though public leaderboards remain led by Anthropic’s Claude Fable 5.1 and Opus 5 at 64.5–65%. The market reflects trader focus on whether the latest Grok release will promptly appear on the official HLE leaderboard at frontier-comparable accuracy, given xAI’s pattern of benchmark emphasis and the benchmark’s rapid score climb from single-digit percentages in early 2025 to the mid-40s–60s% range today. Key near-term catalysts include any verified HLE submission, a potential Grok 4.7 follow-up, and ongoing competition in reasoning and expert-knowledge evaluations.

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No".

If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results.

A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify.

The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere.

If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered.

A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market.

The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Khối lượng
$29,129
Ngày kết thúc
Dec 31, 2026
Thị trường mở
Aug 10, 2026, 6:25 PM ET

Nguồn giải quyết

https://agi.safe.ai/

Người giải quyết

0x65070BE91...

Kết quả đề xuất: Yes

Không tranh chấp

Kết quả cuối cùng: Yes

Cẩn thận với liên kết bên ngoài.

Câu hỏi thường gặp

"Next Grok Model (4.6+): Humanity’s Last Exam Debut?" là thị trường dự đoán trên Polymarket với 5 kết quả có thể nơi các nhà giao dịch mua và bán cổ phần dựa trên điều họ tin sẽ xảy ra. Kết quả dẫn đầu hiện tại là "35%+" ở mức 100%, tiếp theo là "40%+" ở mức 0%. Giá phản ánh xác suất cộng đồng theo thời gian thực. Ví dụ, cổ phần ở giá 100¢ ngụ ý thị trường tập thể cho rằng có 100% khả năng cho kết quả đó. Tỷ lệ này thay đổi liên tục khi trader phản ứng với diễn biến và thông tin mới. Cổ phần đúng kết quả có thể đổi lấy $1 mỗi cổ phần khi thị trường được giải quyết.

Tính đến hôm nay, "Next Grok Model (4.6+): Humanity’s Last Exam Debut?" đã tạo $29.1K tổng khối lượng giao dịch kể từ khi thị trường mở vào Aug 11, 2026. Mức hoạt động giao dịch này phản ánh sự tham gia mạnh mẽ từ cộng đồng Polymarket và giúp đảm bảo tỷ lệ hiện tại được thông tin bởi nhóm người tham gia thị trường sâu rộng. Bạn có thể theo dõi biến động giá trực tiếp và giao dịch trên bất kỳ kết quả nào ngay trên trang này.

Để giao dịch trên "Next Grok Model (4.6+): Humanity’s Last Exam Debut?," duyệt 5 kết quả có sẵn trên trang này. Mỗi kết quả hiển thị giá hiện tại đại diện cho xác suất ngụ ý của thị trường. Để mở vị thế, chọn kết quả bạn tin là có khả năng nhất, chọn "Có" để giao dịch ủng hộ hoặc "Không" để giao dịch chống, nhập số tiền và nhấn "Giao dịch." Nếu kết quả bạn chọn đúng khi thị trường giải quyết, cổ phần "Có" của bạn trả $1 mỗi cổ phần. Nếu sai, chúng trả $0. Bạn cũng có thể bán cổ phần bất cứ lúc nào trước khi giải quyết nếu muốn chốt lời hoặc cắt lỗ.

Ứng viên dẫn đầu hiện tại cho "Next Grok Model (4.6+): Humanity’s Last Exam Debut?" là "35%+" ở mức 100%, nghĩa là thị trường cho 100% khả năng cho kết quả đó. Kết quả gần nhất tiếp theo là "40%+" ở mức 0%. Tỷ lệ cập nhật theo thời gian thực khi trader mua và bán cổ phần, phản ánh cái nhìn tập thể mới nhất về điều có khả năng xảy ra nhất. Kiểm tra thường xuyên hoặc đánh dấu trang này để theo dõi tỷ lệ thay đổi khi thông tin mới xuất hiện.

Quy tắc giải quyết cho "Next Grok Model (4.6+): Humanity’s Last Exam Debut?" định nghĩa chính xác điều gì cần xảy ra để mỗi kết quả được tuyên bố thắng — bao gồm nguồn dữ liệu chính thức được sử dụng để xác định kết quả. Bạn có thể xem tiêu chí giải quyết đầy đủ trong phần "Quy tắc" trên trang này phía trên bình luận. Chúng tôi khuyên đọc kỹ quy tắc trước khi giao dịch, vì chúng chỉ rõ điều kiện, trường hợp ngoại lệ và nguồn chính xác quản lý cách thị trường được thanh toán.