AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
13 Aug 2026

Conflict and Congruency Effects in Large Language Models: In-Weight and In-Context Competition in a Verbal Conflict Task

Model ReleasesDGX agent

arXiv:2608.11510v1 Announce Type: cross Abstract: Congruency effects, observed in conflict tasks such as Stroop and flanker tasks, have been investigated for nearly a century in psychology and neurosc

Dynamics Models for Offline Hyperparameter Selection in Real-World RL

ApplicationsDGX agent

arXiv:2608.11349v1 Announce Type: cross Abstract: A key obstacle to deploying reinforcement learning in real-world systems is hyperparameter selection, particularly when simulators are unavailable and

Evaluating and Calibrating Diffusion Model-derived Uncertainty for Quantitative MRI Mapping

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.11942v1 Announce Type: new Abstract: Quantitative MRI (qMRI) provides standardised tissue parameter maps, but the reliability of deep learning-based qMRI mapping methods is often not explic

ExRole: From Team Trajectories to Executable Roles in Multi-Agent Language Models

Model ReleasesDGX agent

arXiv:2608.11949v1 Announce Type: new Abstract: Roles provide an interpretable interface for organizing language-model agents, yet most multi-agent systems treat them as hand-written prompt labels dis

Powered by @Cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where e…

Model ReleasesDGX agent

Powered by @Cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second counts. Ultrafast is designed for businesses whe

Program Semantic Inequivalence Game with Large Language Models

Model ReleasesDGX agent

arXiv:2505.03818v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) can achieve strong performance on everyday coding tasks, but they can fail on complex tasks that require non-triv

Thanks. The fact that it works so well for ARC is also a clue to the extent of generalization in these models. It was not at all obvious tha…

ResearchDGX agent

Thanks. The fact that it works so well for ARC is also a clue to the extent of generalization in these models. It was not at all obvious that it would work originally because the model is never traine

The Coachella of robotics, world models, and physical AI is Runway’s SF Summit, coming to the Bay Area on September 30. Speaker lineup: @Qua…

HardwareDGX agent

The Coachella of robotics, world models, and physical AI is Runway’s SF Summit, coming to the Bay Area on September 30. Speaker lineup: @QuanVng from @physical_int, @liu_mingyu from @nvidia, @robertni

Understanding Why Foundation Models Work for Diffusion-Generated Image Detection

Local AiDGX agent

arXiv:2608.12155v1 Announce Type: new Abstract: Vision foundation models have recently emerged as powerful feature extractors for detecting AI-generated images, achieving strong generalization across

XYZFlow:Scaling Multi dimensional Shortcut Flows for Efficient Generative Modeling

ResearchDGX agent

arXiv:2608.12276v1 Announce Type: new Abstract: High-fidelity image generation faces a trade-off between speed and quality. Diffusion models produce strong visuals but require costly iterative samplin

12 Aug 2026

AIFS-TC: A simple correction competitive with the operational frontier for tropical cyclone intensity forecasting

Model ReleasesDGX agent

arXiv:2608.09959v1 Announce Type: cross Abstract: AI weather models are in the process of revolutionising weather forecasting. While these models have been shown to achieve superior performance to phy

Can Computational Reducibility Lead to Transferable Models for Graph Combinatorial Optimization?

ResearchDGX agent

arXiv:2603.02462v2 Announce Type: replace-cross Abstract: A key challenge in developing unified neural solvers for combinatorial optimization (CO) is the efficient generalization of models from a give

DSAR: Dual-Stream Autoregressive Modeling of Temporal Cloth Dynamics for Photorealistic Animatable Avatars

TutorialsDGX agent

arXiv:2608.10500v1 Announce Type: new Abstract: Creating photorealistic and temporally coherent animatable human avatars from RGB videos remains challenging. Current methods struggle to capture realis

From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026

Model ReleasesDGX agent

Enterprise content is no longer just something people read. AI apps and agents are only as useful as the information they can understand, yet much of the world’s enterprise knowledge is locked in docu

Frozen Brain-MRI Foundation Models Are Site Fingerprints

ResearchDGX agent

arXiv:2608.10295v1 Announce Type: cross Abstract: Frozen foundation-model (FM) embeddings are increasingly used as off-the-shelf brain-MRI representations, on the assumption that they capture anatomy.

Generation-Step-Aware Framework for Cross-Modal Representation and Control in Multilingual Speech-Text Models

SafetyDGX agent

arXiv:2601.17387v3 Announce Type: replace Abstract: Multilingual speech-text models rely on cross-modal language alignment to transfer knowledge between speech and text, but it remains unclear whether

Google DeepMind launches SL2T, a multilingual sign-language-to-text model debuting on the Pixel 11 in Gboard and Live Transcribe, first with ASL and English (Mike Wheatley/SiliconANGLE)

Model ReleasesDGX agent

Mike Wheatley / SiliconANGLE: Google DeepMind launches SL2T, a multilingual sign-language-to-text model debuting on the Pixel 11 in Gboard and Live Transcribe, first with ASL and English — Google Deep

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Cla…

Model ReleasesDGX agent

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Claude Opus 5 Max Agentic AI is about more than answering quest

IADD-TR: Intervention-Aware Dynamics Decoupling with Targeted Regularization for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2608.10634v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL), which learns environment dynamics to generate synthetic experience, is a promising approach to sample-efficie

Intrinsic Structure: Spectral Identifiability for Mechanistic Interpretability

Model ReleasesDGX agent

arXiv:2608.10172v1 Announce Type: new Abstract: Mechanistic interpretability explains models by identifying circuits inside them, but has no way to tell whether a circuit is a property of the model or

Learning in ImaginationLand: Omnidirectional Policies through 3D Generative Models (OP-Gen)

SafetyDGX agent

arXiv:2509.06191v2 Announce Type: replace-cross Abstract: Recent 3D generative models, which are capable of generating full object shapes from just a few images, now open up new opportunities in robot

Link-adaptive digital twin for robust physical-layer modeling in hybrid-amplified ultra-wideband optical networks

TutorialsDGX agent

arXiv:2608.10517v1 Announce Type: cross Abstract: Accurate physical-layer modeling is increasingly essential for reliable ultra-wideband operation and capacity optimization, especially under the inten

Long-Time Trajectory Approximation via SA-NODEs: Model Predictive and Floquet Strategies

Model ReleasesDGX agent

arXiv:2608.10738v1 Announce Type: new Abstract: We study the approximation of dynamical systems by semi-autonomous neural ordinary differential equations (SA-NODEs) over long time horizons. For a sing

MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models

SafetyDGX agent

arXiv:2608.10635v1 Announce Type: cross Abstract: Medical Vision-Language Models (Med-VLMs) excel at verbalizing visual content, yet precise visual perception, segmentation, and grounding remain chall

Multi-View Relational Distillation for Spatial Reasoning with Vision-Language Models

SafetyDGX agent

arXiv:2608.10864v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong image and video understanding, yet their visual-spatial representations remain geometrically fragile,

P3CA: Encoder-Agnostic Interpretation of Vision Foundation Model Embeddings via Spatial Probing

Local AiDGX agent

arXiv:2608.10131v1 Announce Type: new Abstract: Vision foundation models are increasingly used as reusable encoders in medical image computing, yet their high-dimensional spatial embeddings are diffic

Post-Hoc Sparse Coding of Latent Communication Between Vision-Language Model Agents

ResearchDGX agent

arXiv:2608.10198v1 Announce Type: new Abstract: Latent-space communication allows heterogeneous vision-language model agents to exchange continuous representations without serializing visual and reaso

Pretrained Optimization Model for Zero-Shot Black Box Optimization

Model ReleasesDGX agent

arXiv:2405.03728v3 Announce Type: replace-cross Abstract: Zero-shot optimization involves optimizing a target task that was not seen during training, aiming to provide the optimal solution without or

Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy Datasets

ApplicationsDGX agent

arXiv:2608.10657v1 Announce Type: cross Abstract: Leukemia cell image classification is challenged by real-world domain shifts from acquisition, staining, illumination, and site protocols, causing sin

SynBoost: A Synergistic Framework for Fast Sampling of Diffusion Models

ResearchDGX agent

arXiv:2506.13058v2 Announce Type: replace-cross Abstract: Diffusion probabilistic models (DPMs) have demonstrated remarkable success in visual generation. However, their iterative sampling mechanism r

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and…

Model ReleasesDGX agent

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and delivers particularly strong results on document understand

11 Aug 2026

Automated Generation of Complexity-Validated Decision Scenarios Using Large Language Models

Model ReleasesDGX agent

arXiv:2608.08822v1 Announce Type: new Abstract: Cognitive decision-making research depends on diverse scenarios with carefully controlled complexity, yet manual production is slow, inconsistent, and b

b10361

Model ReleasesDGX agent

model : fix SWA not being enabled for EXAONE 4.5 (#26848) model : fix SWA not being enabled for EXAONE 4.5 load_arch_hparams tests hparams.n_layer() == 64 before LLM_KV_NEXTN_PREDICT_LAYERS has been r

Cognitive Energy Modeling for Neuroadaptive Human-Machine Systems using EEG and WGAN-GP

ResearchDGX agent

arXiv:2604.01653v2 Announce Type: replace Abstract: Electroencephalography (EEG) provides a non-invasive insight into the brain's cognitive and emotional dynamics. However, modeling how these states e

Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models

SafetyDGX agent

arXiv:2608.08829v1 Announce Type: cross Abstract: Activation steering edits the behaviour of a frozen language model by adding a learned vector to its residual stream, and current practice fixes the i

FanarGuard: A Culturally-Aware Moderation Filter for Arabic Language Models

Model ReleasesDGX agent

arXiv:2511.18852v2 Announce Type: replace Abstract: Content moderation filters are a critical safeguard against alignment failures in language models. Yet most existing filters focus narrowly on gener

FlowErase-OPD: Multi-Concept Erasure via Anchored On-Policy Distillation in Flow Matching Models

SafetyDGX agent

arXiv:2608.07620v1 Announce Type: new Abstract: Recent advances in flow matching models have substantially improved the quality of text-to-image generation, but have also raised increasing safety conc

From Uncertainty to Failure Attribution: Self-Diagnosing Models for Failure Attribution under Distribution Shift

ResearchDGX agent

arXiv:2608.07953v1 Announce Type: new Abstract: Distribution shift poses a significant challenge to the robustness of machine learning models, but the current solutions only aim to detect out-of-distr

How sensitive do we want AI to be? Socio-communicative competencies of large language models in healthcare

Model ReleasesDGX agent

arXiv:2608.07511v1 Announce Type: cross Abstract: Background. Effective clinical practice relies heavily on the socio-communicative skills of medical professionals. Large language models (LLMs) have b

How to Ask the AI: A User Perspective Survey for Large Language Model Prompting

Model ReleasesDGX agent

arXiv:2608.07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by t

I tested the CMP170HX

Model ReleasesDGX agent

Lots of rumor and misinfo bouncing around, so I put some of these old mining cards to the test. I used 4 of the 8GB cards, set to 64GB each. Lots of models fit entirely on a single card, and you can a

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models a…

Model ReleasesDGX agent

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing the frontier of coding and knowledge work, but su

Jagle: Building a Large-Scale Japanese Multimodal Post-Training Dataset for Vision-Language Models

ResearchDGX agent

arXiv:2604.02048v2 Announce Type: replace Abstract: Developing vision-language models (VLMs) that generalize across diverse tasks requires large-scale training datasets with diverse content. In Englis

JEPA-WAM: Learning Vision-Language-Action Policies with Joint-Embedding World Modeling

SafetyDGX agent

arXiv:2608.09381v1 Announce Type: new Abstract: Robust robot control benefits from explicitly modeling state transitions, but video-generation world action models (WAMs) introduce substantial deployme

Large Language Models Align with the Human Brain during Creative Thinking

Model ReleasesDGX agent

arXiv:2604.03480v2 Announce Type: replace-cross Abstract: Creative thinking is a fundamental aspect of human cognition, and divergent thinking-the capacity to generate novel and varied ideas-is widely

Multi-modal Interactive Control of Robotic Arm based on Offline Large Language Models

Local AiDGX agent

arXiv:2608.08183v1 Announce Type: new Abstract: Large Language Models (LLMs) have significantly revolutionized the modern society with numerous advanced interactions between humans and AI agents, wher

Population-Level Generative Modeling for Ranking Data

Model ReleasesDGX agent

arXiv:2608.08422v1 Announce Type: cross Abstract: Ranking data arise in scientific and machine learning applications, including recommendation systems, information retrieval, voting, marketing, and AI

Protecting patient privacy in clinical foundation models: Technical and legal perspectives

ApplicationsDGX agent

arXiv:2608.07705v1 Announce Type: new Abstract: Clinical foundation models trained on large-scale patient data are increasingly used for decision support, screening, and public health. As deployment e

Router Sensitivity Under Lightweight Fine-Tuning Identifies Prunable Experts in Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2608.07890v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models decouple total parameters from per-token compute, but deployment still requires storing every expert. Recent theory sh

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models

ApplicationsDGX agent

arXiv:2608.08839v1 Announce Type: cross Abstract: World-Action Models (WAMs) have emerged as a promising paradigm for robotic manipulation. However, most existing WAMs generate future videos and actio

SuperCoder: Assembly Program Superoptimization with Large Language Models

Model ReleasesDGX agent

arXiv:2505.11480v4 Announce Type: replace-cross Abstract: Superoptimization is the task of transforming a program into a faster one, and ideally the very fastest possible one, while preserving its inp

The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World

AgentsDGX agent

arXiv:2608.08239v1 Announce Type: cross Abstract: LLM routers promise efficiency by matching each request to the cheapest adequate model, and are increasingly applied per step inside multi-step agents

The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

Model ReleasesDGX agent

arXiv:2608.08654v1 Announce Type: new Abstract: How much an AI coding agent costs to run can depend more on the agent scaffolding that drives it than on the interface through which it reaches its tool

VLZip: Unified Visual and Textual Compression for Interleaved Long-Context Modeling

Model ReleasesDGX agent

arXiv:2608.08630v1 Announce Type: new Abstract: Vision Language Models (VLMs) face significant challenges with ultra-long, interleaved image-text sequences due to the quadratic complexity of self-atte

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

Model ReleasesDGX agent

arXiv:2608.00591v2 Announce Type: replace Abstract: A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can aris

WorldSimProbe: Diagnosing Simulator Faithfulness in Action-Conditioned World Models for Embodied Manipulation

Model ReleasesDGX agent

arXiv:2608.09298v1 Announce Type: cross Abstract: Action-conditioned world models (ACWMs) promise to provide embodied AI with scalable predictive simulators for planning, policy evaluation, and data g

10 Aug 2026

Anthropic details an unreleased Claude model's attempt to solve the Riemann hypothesis; it didn't solve it but 'unexpectedly' made strides on a related problem (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic details an unreleased Claude model's attempt to solve the Riemann hypothesis; it didn't solve it but “unexpectedly” made strides on a related problem — Recently, a member of staff

Beyond Myopic World Models: Long-Horizon End-to-End Training for Direct Future Prediction

Local AiDGX agent

arXiv:2608.07420v1 Announce Type: new Abstract: World models are expected to support imagination over extended temporal horizons, yet most are still trained through local few-step prediction objective

Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Reward Models via Contribution Contrast

ResearchDGX agent

arXiv:2608.06400v1 Announce Type: new Abstract: Reward models are central to learning from human preferences, yet identifying what drives their predictions remains challenging. Recent sparse Mixture-o

Chat UIs with native audio input for multimodal models?

Model ReleasesDGX agent

I've been running Gemma 4 E4B with oMLX and I can't find any chat interfaces that directly send the audio file to the model instead of running the audio through a separate STT layer. I can confirm the

← Previous
1…979899100101…1009
Next →