Skip to content
GCC AI Research

Search

Results for "reasoning"

UAE’s Falcon 40B Dominates Leaderboard: Ranks #1 Globally in Latest Hugging Face Independent Verification of Open-source AI Models

TII ·

TII's Falcon 40B, a 40-billion-parameter open-source AI model, has ranked #1 on Hugging Face's Open LLM Leaderboard, surpassing models like LLaMA and StableLM. The leaderboard uses benchmarks like AI2 Reasoning Challenge, HellaSwag, MMLU, and TruthfulQA. Trained on one trillion tokens, Falcon 40B's weights are available for research and commercial use. Why it matters: This achievement positions the UAE as a leader in generative AI and promotes transparent, inclusive AI development.

In recognition of Sheikh Khalifa’s contribution to advancing science and technology, UAE President endorses launch of K2 Think, world’s most advanced open-source reasoning model - wam.ae

WAM ·

The UAE President has endorsed the launch of K2 Think, which is described as the world’s most advanced open-source reasoning model. This launch recognizes Sheikh Khalifa’s contributions to advancing science and technology within the UAE. The announcement signifies a major national initiative in the field of artificial intelligence development. Why it matters: This positions the UAE at the forefront of open-source AI innovation and advanced reasoning capabilities, potentially setting new benchmarks for global AI development.

Saudi Arabia and NVIDIA to Build AI Factories to Power Next Wave of Intelligence for the Age of Reasoning - NVIDIA Newsroom

SDAIA ·

Saudi Arabia is collaborating with NVIDIA to develop and build 'AI factories' within the Kingdom. These 'AI factories' will accelerate the development and deployment of generative AI and other advanced AI applications, providing powerful computing infrastructure. The initiative aims to support Saudi Arabia's vision of becoming a global leader in AI development, enabling what NVIDIA terms the 'Age of Reasoning.' Why it matters: This major strategic partnership signifies Saudi Arabia's significant investment in advanced AI infrastructure, positioning the Kingdom as a key player in the global AI landscape and fostering domestic AI innovation.

Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

arXiv ·

Researchers have introduced VISE (Visual Invariance Self-Evolution), a purely unsupervised framework designed to address 'visual under-conditioning' in self-evolving Large Multimodal Models (LMMs). VISE utilizes geometric and semantic invariance-based rewards to directly regularize the model's visual conditioning, ensuring it attends to visual content rather than relying on language priors. Trained on raw unlabeled images, experiments using Qwen3-VL-2B demonstrate significant performance gains, including +16.85 CIDEr on COCO and a 5.0-point reduction in object hallucination across 18 benchmarks. Why it matters: This research from MBZUAI offers a significant advancement in improving the visual reasoning capabilities and reliability of LMMs in unsupervised settings, making them more robust for real-world applications.

Yasi One launches from Abu Dhabi: A new AI that thinks before it responds - Gulf News

Gulf News ·

Yasi One, a new artificial intelligence system, has been launched from Abu Dhabi, UAE. This AI is specifically noted for its unique capability to 'think before it responds,' suggesting advanced processing and reasoning functionalities. The launch of Yasi One was reported by Gulf News. Why it matters: This development underscores Abu Dhabi's growing ambition to develop and deploy cutting-edge AI technologies, potentially contributing to more sophisticated and contextually aware AI applications in the region.

World Reasoning Arena

arXiv ·

Researchers from MBZUAI have introduced WR-Arena, a new comprehensive benchmark designed to evaluate World Models (WMs) beyond traditional next-state prediction and visual fidelity. WR-Arena assesses WMs across three core dimensions: Action Simulation Fidelity, Long-horizon Forecast, and Simulative Reasoning and Planning, using a curated task taxonomy and diverse datasets. Extensive experiments with state-of-the-art WMs revealed a significant gap between current models' capabilities and human-level hypothetical reasoning. Why it matters: This benchmark provides a critical diagnostic tool and guideline for developing more robust and intelligent world models capable of advanced understanding, forecasting, and purposeful action, particularly for AI research in the region.

CoVR-R:Reason-Aware Composed Video Retrieval

arXiv ·

A new approach to composed video retrieval (CoVR) is presented, which leverages large multimodal models to infer causal and temporal consequences implied by an edit. The method aligns reasoned queries to candidate videos without task-specific finetuning. A new benchmark, CoVR-Reason, is introduced to evaluate reasoning in CoVR.

Safran and the Technology Innovation Institute intend to lead the next evolution in geospatial intelligence

TII ·

Safran.AI and the Technology Innovation Institute (TII) intend to form a strategic alliance to develop a next-generation Agentic AI geospatial intelligence (GEOINT) platform. The platform will combine Safran.AI’s GEOINT expertise with TII’s expertise in Agentic AI and orchestration platforms, enabling autonomous reasoning and transforming spaceborne imagery into decision-grade intelligence. The collaboration will focus on three major technological streams. Why it matters: This partnership signifies a major advancement in sovereign geospatial intelligence capabilities within the UAE, moving from traditional analysis to autonomous understanding for enhanced national security and decision-making.