Skip to content
GCC AI Research

Search

Results for "large language models"

UAE gets easier access to US AI chips: What changed and why it matters - Gulf News

Gulf News ·

The UAE has reportedly gained easier access to high-performance artificial intelligence (AI) chips from the United States, following recent adjustments in US export regulations. This development is expected to streamline the procurement process for advanced AI hardware, which is crucial for developing large language models and other compute-intensive AI applications. The specific policy changes and licensing requirements facilitating this access will significantly impact technology companies and research institutions operating within the UAE. Why it matters: This enhanced access to critical AI infrastructure is vital for accelerating the UAE's national AI strategy and strengthening its position as a global AI hub.

UAE receives first shipment of Nvidia's advanced AI chips - The National

The National ·

The UAE has received its initial shipment of advanced AI chips from Nvidia, marking a significant milestone in its national AI strategy. These chips are essential for powering the country's growing supercomputing capabilities and accelerating the development of large language models. This delivery underscores the UAE's commitment to establishing itself as a global leader in AI innovation. Why it matters: This acquisition directly enhances the UAE's capacity for advanced AI research and development, solidifying its competitive position in the global AI landscape and fostering local technological growth.

G42 & R/GA Launch Alpha.G42.ai: A World-First Generative Interface, Prototyping the Future of the Web

G42 ·

G42, a global leader in artificial intelligence based in Abu Dhabi, partnered with creative innovation company R/GA to launch alpha.G42.ai, a generative interface designed to transform traditional websites into dynamic, conversational systems. This prototype redefines a brand's digital presence by employing an intelligent agent powered by integrated large language models (LLMs) to generate and curate personalized content for each visitor in real-time. The system processes various content types as knowledge, which it then synthesizes to produce dynamic, tailored outputs for users interacting via voice or text, moving beyond static content management. Why it matters: This initiative from a major UAE AI firm pioneers a novel approach to web interfaces, potentially influencing future digital interactions and content delivery globally.

HalluTruthQA-4K: A Fine-Grained Corpus and Annotation Process for Arabic Hallucination Detection and Truth Verification

arXiv ·

Researchers introduced HalluTruthQA-4K, an expanded corpus comprising 4,000 expert-curated Arabic question-answering instances designed for hallucination detection and truth verification. This resource spans four knowledge-intensive domains: Islamic knowledge, history, science, and geography, and serves as the official dataset for Track 2 of the HalluScoring 2026 shared task. For hallucinated responses, the corpus provides character-level erroneous spans, human-written explanations, and hierarchical hallucination types, alongside verified reference answers and distractors. Why it matters: HalluTruthQA-4K provides a crucial fine-grained resource for evaluating and improving the factual reliability and trustworthiness of Arabic large language models.

Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs

arXiv ·

This study investigates methods to steer Arabic Large Language Models (LLMs) towards generating specific dialects, addressing the challenge of data scarcity for dialectal Arabic. Researchers identified sparse neuron populations encoding dialect-specific features and developed a vector-steering approach using dialect-specific activation directions. These inference-time methods allow for controlling dialectal output by amplifying or suppressing neuron activity or injecting specific vectors. Why it matters: This research offers a principled, interpretability-grounded framework to improve dialectal accuracy in Arabic LLMs without fine-tuning, crucial for enhancing their utility in the diverse Arabic-speaking world.

The Cylindrical Representation Hypothesis for Language Model Steering

arXiv ·

Researchers have proposed the Cylindrical Representation Hypothesis (CRH) to address the instability and unpredictability observed in steering large language models, an issue not fully explained by the existing Linear Representation Hypothesis (LRH). CRH suggests that overlapping concept contributions lead to a sample-specific axis-orthogonal structure, comprising a central axis for concept generation and a surrounding normal plane for steering sensitivity. This framework identifies intrinsic uncertainty at the 'sensitive sector' level within the plane, providing a principled explanation for fluctuations in steering outcomes. Experiments verify the existence of this cylindrical structure and demonstrate CRH's practical utility in interpreting real-world model steering behavior, with code available on GitHub from mbzuai-nlp. Why it matters: This research from MBZUAI offers a crucial theoretical advancement in understanding and potentially improving the control and reliability of large language models.

The Cylindrical Representation Hypothesis for Language Model Steering

arXiv ·

Researchers from MBZUAI have proposed the Cylindrical Representation Hypothesis (CRH) to explain the instability and unpredictability observed in large language model steering. CRH relaxes the orthogonality assumption of the existing Linear Representation Hypothesis, positing a cylindrical structure where a central axis captures concept differences and a surrounding normal plane controls steering sensitivity. The hypothesis suggests that the intrinsic uncertainty in identifying specific sensitive sectors within this normal plane accounts for why steering outcomes frequently fluctuate even with well-aligned directions. Why it matters: This research offers a more robust theoretical framework for understanding and potentially improving the control and reliability of large language models.

Instruction-Guided Poetry Generation in Arabic and Its Dialects

arXiv ·

Researchers at MBZUAI have developed a new method for controllable poetry generation in Arabic and its dialects, moving beyond traditional analysis tasks for Arabic poetry within Large Language Models (LLMs). They introduce a large-scale, instruction-based dataset in Modern Standard Arabic (MSA) and various Arabic dialects, enabling LLMs to perform tasks like writing, revising, and continuing poems based on user criteria. Experiments show that fine-tuning LLMs on this dataset results in models capable of generating poetry aligned with user requirements, validated by automated metrics and human evaluation. Why it matters: This work represents a significant advancement in Arabic Natural Language Processing, offering tools for creative expression and cultural preservation while opening new avenues for user-guided content generation in culturally rich text forms.