{"id":764,"date":"2025-11-18T16:41:31","date_gmt":"2025-11-18T08:41:31","guid":{"rendered":"https:\/\/www.chain258.com\/?p=764"},"modified":"2025-11-18T16:41:31","modified_gmt":"2025-11-18T08:41:31","slug":"xai-has-officially-released-grok-4-1","status":"publish","type":"post","link":"https:\/\/www.chain258.com\/index.php\/2025\/11\/18\/xai-has-officially-released-grok-4-1\/","title":{"rendered":"xAI has officially released Grok 4.1"},"content":{"rendered":"\n<figure class=\"wp-block-image size-full\"><img loading=\"lazy\" decoding=\"async\" width=\"960\" height=\"658\" src=\"https:\/\/www.chain258.com\/wp-content\/uploads\/2025\/11\/676d893bb64c04318e587878f8891d5f_interlace1.jpg\" alt=\"\" class=\"wp-image-765\" srcset=\"https:\/\/www.chain258.com\/wp-content\/uploads\/2025\/11\/676d893bb64c04318e587878f8891d5f_interlace1.jpg 960w, https:\/\/www.chain258.com\/wp-content\/uploads\/2025\/11\/676d893bb64c04318e587878f8891d5f_interlace1-300x206.jpg 300w, https:\/\/www.chain258.com\/wp-content\/uploads\/2025\/11\/676d893bb64c04318e587878f8891d5f_interlace1-768x526.jpg 768w\" sizes=\"auto, (max-width: 960px) 100vw, 960px\" \/><\/figure>\n\n\n\n<p class=\"wp-block-paragraph\">xAI has officially released Grok 4.1. This version is now available to all users on grok.com, the X platform, as well as iOS and Android apps, including free users, and is enabled by default in Auto mode.<br>Elon Musk, founder of xAI, stated that users will &#8220;noticeably feel improvements in speed and quality.&#8221; Unlike previous updates that focused on computing power or scale, Grok 4.1 emphasizes three intuitive yet highly challenging directions: faster responses, higher factual accuracy, and a more natural and personalized conversational experience.<br>Performance Improvements: Fewer Hallucinations, Higher Fact Accuracy, Stronger Style Control<br>Grok 4.1 performs exceptionally well in information query tests. Official data shows: Grok 4.1&#8217;s hallucination rate has dropped from 12.09% to 4.22%, a reduction of nearly threefold; FActScore has improved from 9.89% to 2.97%, also showing significant enhancement. Against the backdrop of widespread factual instability in current large models, this represents a genuine structural upgrade.<br>xAI stated that the performance improvement of Grok 4.1 is due to its reinforcement learning infrastructure and new reward model system: Grok 4.1 uses a &#8220;cutting-edge reasoning model&#8221; as the reward model, enabling the model to self-evaluate and iterate quickly. This means training is no longer overly reliant on large-scale manual annotation, and also makes style, tone, and collaborative capabilities more controllable.<br>Grok 4.1 achieved a blind preference rate of 64.78% in silent testing<br>In the most recent round of silent testing (from November 1 to 14), Grok 4.1 achieved a blind preference rate of 64.78%, significantly higher than the previous version.<br>Grok 4.1&#8217;s performance on the LMSYS Arena<br>Grok 4.1 has shown a leap in performance on the international blind testing platform LMSYS Arena. In the latest evaluation round, Grok 4.1&#8217;s Thinking mode (code-named quasarflux) achieved an Elo rating of 1483 (Elo rating system, used to measure the relative strength of models in blind test battles), ranking first among all publicly available models; its non-reasoning mode also reached 1465 Elo, ranking second. This achievement is rare in itself\u2014it outperforms many other models that use full reasoning configurations, even without employing a chain-of-thought approach.<br>In comparison, the previous generation Grok 4 was ranked 33rd overall. Now, Grok 4.1 has not only jumped a rank but also signifies that its foundational conversational quality and comprehensive capabilities have steadily entered the industry&#8217;s top tier.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>xAI has officially released Gr&hellip;<\/p>\n","protected":false},"author":2,"featured_media":765,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[9,8],"tags":[24,26,16],"class_list":["post-764","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-companies","category-deep-tech","tag-ai","tag-deep-tech","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/posts\/764","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/comments?post=764"}],"version-history":[{"count":1,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/posts\/764\/revisions"}],"predecessor-version":[{"id":766,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/posts\/764\/revisions\/766"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/media\/765"}],"wp:attachment":[{"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/media?parent=764"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/categories?post=764"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.chain258.com\/index.php\/wp-json\/wp\/v2\/tags?post=764"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}