AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
19 May 2026

Automatic Unsupervised Ensemble Outlier Model Selection--Extended Version

ApplicationsDGX agent

arXiv:2605.16567v1 Announce Type: cross Abstract: Unsupervised outlier detection is attractive because it eliminates the need for labeled data. Moreover, forming multi-model ensembles can improve dete

Beyond Accuracy: Robustness, Interpretability and Expressiveness of EEG Foundation Models

ResearchDGX agent

arXiv:2605.17562v1 Announce Type: cross Abstract: EEG foundation models (EEG-FMs) have been evaluated predominantly on clean, in-distribution accuracy, leaving their robustness, interpretability and r

Beyond Neural Incompatibility: Cross-Scale Knowledge Transfer in Language Models through Latent Semantic Alignment

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.24208v2 Announce Type: replace Abstract: Language Models (LMs) encode substantial knowledge in their parameters, yet it remains unclear how to transfer such knowledge in a fine-grained mann

DARE-EEG: A Foundation Model for Mining Dual-Aligned Representation of EEG

Model ReleasesDGX agent

arXiv:2605.18298v1 Announce Type: new Abstract: Foundation models pre-trained through masked reconstruction on large-scale EEG data have emerged as a promising paradigm for learning generalizable neur

Democratizing Large-Scale Re-Optimization with LLM-Guided Model Patches

AgentsDGX agent

arXiv:2605.18692v1 Announce Type: new Abstract: Optimization models developed by operations research (OR) experts are often deployed as decision-support systems in industrial settings. However, real-w

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games?

Model ReleasesDGX agent

arXiv:2605.11223v2 Announce Type: replace Abstract: Vision-Language(-Action) Models (VLMs) are increasingly applied to interactive environments, yet existing benchmarks often overlook the complex phys

DP-SelFT: Differentially Private Selective Fine-Tuning for Large Language Models

Model ReleasesDGX agent

arXiv:2605.17432v1 Announce Type: new Abstract: Large language models (LLMs) are commonly adapted to downstream tasks through fine-tuning, but fine-tuning data often contains sensitive information tha

Effort as Ceiling, Not Dial: Reasoning Budget Does Not Modulate Cognitive Cost Alignment Between Humans and Large Reasoning Models

Model ReleasesDGX agent

arXiv:2605.16938v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) generate chain-of-thought traces whose length tracks human reaction times across cognitive tasks, but recent debate ques

Emulating the Forced Response of Climate Models with Flow Matching

ResearchDGX agent

arXiv:2605.16929v1 Announce Type: new Abstract: Global climate models are essential tools to simulate past and potential future pathways of climate change, as well as associated climate impacts. Share

Evaluating AI Alignment in LLMs: Output Analysis of Value Priorities Across 75 Models with Human Benchmarking

SafetyDGX agent

arXiv:2506.12617v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in human-AI interaction research and practice, yet existing capability and safety benchmarks reve

Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models

Local AiDGX agent

arXiv:2604.01404v2 Announce Type: replace-cross Abstract: How do language models retrieve entity-specific facts from their parameters? We investigate this question by searching for sparse, entity-sele

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

SafetyDGX agent

arXiv:2512.23180v3 Announce Type: replace Abstract: Driving World Models (DWMs) have been developing rapidly with the advances of generative models. However, existing DWMs lack 3D scene understanding

Google launches new Omni model. Google says that today Omni can create and edit (yes, edit!!) video but eventually it is to go from “any inp…

Model ReleasesDGX agent

Google launches new Omni model. Google says that today Omni can create and edit (yes, edit!!) video but eventually it is to go from “any input to any output”. Even more of a reason to believe that Sor

HalluScore: Large Language Model Hallucination Question Answering Benchmark

Model ReleasesDGX agent

arXiv:2605.17007v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to g

Incantation: Natural Language as the Action Interface for Multi-Entity Video World Models

Model ReleasesDGX agent

arXiv:2605.18601v1 Announce Type: new Abstract: Modern interactive video world models have achieved impressive visual fidelity, yet lack fine-grained multi-entity control and cross-entity, cross-world

LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models

Model ReleasesDGX agent

arXiv:2605.17653v1 Announce Type: cross Abstract: Sub-billion-parameter Transformer language models are increasingly deployed on edge devices, where the privacy, latency, and operating-cost advantages

MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models

Model ReleasesDGX agent

arXiv:2602.12871v2 Announce Type: replace Abstract: Large language models (LLMs) have attracted growing interest as supportive tools for psychiatric assessment and clinical decision support. However,

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

Model ReleasesDGX agent

arXiv:2605.17198v1 Announce Type: cross Abstract: To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to

Scale Determines Whether Language Models Organize Representation Geometry for Prediction

SafetyDGX agent

arXiv:2605.17084v1 Announce Type: cross Abstract: In language models, what a representation encodes is determined by the geometry of its representation space: distances, not activations, carry meaning

Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency

SafetyDGX agent

arXiv:2605.18162v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have made striking progress, yet their spatial reasoning remains fragile: models that answer an original input correctly

Self-Supervised On-Policy Distillation for Reasoning Language Models

Model ReleasesDGX agent

arXiv:2605.17497v1 Announce Type: new Abstract: GRPO-style RLVR trains reasoning models from multiple on-policy attempts per prompt, but typically uses these attempts only through terminal rewards. We

SWoMo: Neuro-Symbolic World Model for Cataract Surgery Simulation

AgentsDGX agent

arXiv:2605.16530v1 Announce Type: new Abstract: Realistic surgical simulation plays a crucial role in training novice surgeons and in the development of autonomous agents. World models can scale such

Symmetry Matters: Auditing and Symmetrizing 3D Generative Models

ResearchDGX agent

arXiv:2512.18953v2 Announce Type: replace Abstract: Symmetry is a strong prior present in many object categories, yet standard benchmarks for 3D generative models rarely report whether this prior is p

Task Abstention for Large Language Models in Code Generation

Model ReleasesDGX agent

arXiv:2605.17029v1 Announce Type: cross Abstract: Large language models (LLMs) have revolutionized automated code generation. One serious concern, however, is the so-called ``hallucination'', i.e., LL

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

Model ReleasesDGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

Time Series Foundation Models as Strong Baselines in Transportation Forecasting: A Large-Scale Benchmark Analysis

Model ReleasesDGX agent

arXiv:2602.24238v2 Announce Type: replace Abstract: Accurate forecasting of transportation dynamics is essential for urban mobility and infrastructure planning. Although recent work has achieved stron

Towards Universal Physical Adversarial Attacks via a Joint Multi-Objective and Multi-Model Optimization Framework

SafetyDGX agent

arXiv:2605.17772v1 Announce Type: new Abstract: Physical adversarial attacks often overfit single surrogate models and optimization objectives. While ensemble attacks can mitigate this, existing metho

Transitivity Meets Cyclicity: Explicit Preference Decomposition for Dynamic Large Language Model Alignment

Model ReleasesDGX agent

arXiv:2605.17342v1 Announce Type: cross Abstract: Standard RLHF relies on transitive scalar rewards, failing to capture the cyclic nature of human preferences. While some approaches like the General P

Universal Inverse Distillation for Matching Models with Real-Data Supervision (No GANs)

ResearchDGX agent

arXiv:2509.22459v4 Announce Type: replace-cross Abstract: While achieving exceptional generative quality, modern diffusion, flow, and other matching models suffer from slow inference, as they require

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?

Model ReleasesDGX agent

arXiv:2602.04802v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have achieved impressive performance in cross-modal understanding across textual and visual inputs, yet existing bench

When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings

Model ReleasesDGX agent

arXiv:2605.16288v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in clinical and care settings. This exploratory study investigates whether LLMs exhibit sycophantic

18 May 2026

A Model Can Help Itself: Reward-Free Self-Training for LLM Reasoning

ResearchDGX agent

arXiv:2510.18814v3 Announce Type: replace-cross Abstract: Can language models improve their reasoning performance without external rewards, using only their own sampled responses for training? We show

AGC: Adaptive Geodesic Correction for Adversarial Robustness on Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.15584v1 Announce Type: new Abstract: Vision-language models like CLIP have demonstrated remarkable zero-shot transfer capabilities. However, their susceptibility to imperceptible adversaria

Brain-OF: An Omnifunctional Foundation Model for fMRI, EEG and MEG

ResearchDGX agent

arXiv:2602.23410v3 Announce Type: replace-cross Abstract: Brain foundation models have achieved remarkable advances across a wide range of neuroscience tasks. However, most existing models are limited

CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion

Model ReleasesDGX agent

arXiv:2601.09512v2 Announce Type: replace-cross Abstract: To teach robots complex manipulation tasks, a common approach is to fine-tune a pre-trained vision-language-action model (VLA) on task-specifi

Entity-Centric World Models: Interaction-Aware Masking for Causal Video Prediction

Model ReleasesDGX agent

arXiv:2605.15466v1 Announce Type: new Abstract: Learning predictive world models from unlabelled video is a foundational challenge in artificial intelligence. While Joint Embedding Predictive Architec

EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy

SafetyDGX agent

arXiv:2605.15711v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities across various tasks, yet they remain vulnerable to backdoor attacks. Exi

Going Beyond the Edge: Distributed Inference of Transformer Models on Ultra-Low-Power Wireless Devices

ApplicationsDGX agent

arXiv:2605.15694v1 Announce Type: new Abstract: Transformer models are rapidly becoming a cornerstone of modern Internet of Things (IoT) applications, yet their computational and memory demands far ex

Improve Large Language Model Systems with User Logs

SafetyDGX agent

arXiv:2602.06470v2 Announce Type: replace-cross Abstract: Scaling training data and model parameters has long driven progress in large language models (LLMs), but this paradigm is increasingly constra

Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance

ResearchDGX agent

arXiv:2605.15436v1 Announce Type: new Abstract: This paper presents a comprehensive analysis of neural activation patterns across six distinct large language model (LLM) architectures, examining their

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models

Model ReleasesDGX agent

arXiv:2509.22739v3 Announce Type: replace-cross Abstract: Language models (LMs) are typically post-trained for desired capabilities and behaviors via weight-based or prompt-based steering, but the for

Privacy Evaluation of Generative Models for Trajectory Generation

ResearchDGX agent

arXiv:2605.15246v1 Announce Type: new Abstract: Trajectory data is fundamental to modern urban intelligence, yet its sensitivity raises significant privacy concerns. Generative models such as Generati

RoiMAM: Region-of-Interest Medical Attention Model for Efficient Vision-Language Understanding

ResearchDGX agent

arXiv:2605.15561v1 Announce Type: new Abstract: Vision-Language Models (VLMs) facilitate medical visual question answering (MedVQA) by jointly interpreting images and text. However, existing models ty

Simultaneous State Estimation and Online Model Learning in a Soft Robotic System

TutorialsDGX agent

arXiv:2602.14092v2 Announce Type: replace-cross Abstract: Operating complex real-world systems, such as soft robots, can benefit from precise predictive control schemes that require accurate state and

Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release (Ashley Gold/Axios)

Model ReleasesDGX agent

Ashley Gold / Axios: Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release — A group of more tha

Toward World Modeling of Physiological Signals with Chaos-Theoretic Balancing and Latent Dynamics

ApplicationsDGX agent

arXiv:2605.15465v1 Announce Type: new Abstract: Physiological time series signals reflect complex, multi-scale dynamical processes of the human body. Existing modeling studies focus on static tasks su

Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

AgentsDGX agent

arXiv:2510.10454v2 Announce Type: replace Abstract: Large language models (LLMs) offer a generalizable approach for modeling patient trajectories, but suffer from the long and noisy nature of electron

Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models

Model ReleasesDGX agent

arXiv:2605.15413v1 Announce Type: new Abstract: Despite the remarkable success of transformer architectures in natural language processing, their scalability limitations remain poorly understood throu

Very cool to see Cursor doubling down on training great models. In my opinion, ultimately all serious companies in AI will want to train mod…

IndustryDGX agent

Very cool to see Cursor doubling down on training great models. In my opinion, ultimately all serious companies in AI will want to train models themselves, based on open-source instead of outsourcing

Video Models Can Reason with Verifiable Rewards

SafetyDGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

17 May 2026

When do you reach for other models instead of Claude? What can we do better? Hit me with all of your frustrations. dms open. If you can give…

Model ReleasesDGX agent

When do you reach for other models instead of Claude? What can we do better? Hit me with all of your frustrations. dms open. If you can give me detail (e.g. specifics/transcipts) - it'll help a lot in

15 May 2026

AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

Model ReleasesDGX agent

arXiv:2509.26100v2 Announce Type: replace Abstract: The rapid integration of Large Language Models (LLMs) into high-stakes domains necessitates reliable safety and compliance evaluation. However, exis

Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model

SafetyDGX agent

arXiv:2605.14723v1 Announce Type: new Abstract: Sepsis management in the ICU requires sequential treatment decisions under rapidly evolving patient physiology. Although large language models (LLMs) en

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

Model ReleasesDGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

Energy-Regularized Sequential Model Editing on Hyperspheres

ApplicationsDGX agent

arXiv:2510.01172v3 Announce Type: replace Abstract: Large language models (LLMs) require constant updates to remain aligned with evolving real-world knowledge. Model editing offers a lightweight alter

Exploring Vision-Language Models for Online Signature Verification: A Zero-Shot Capability Study

Model ReleasesDGX agent

arXiv:2605.14845v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have demonstrated strong capabilities in general visual reasoning, yet their applicability to rigor

Generative Bayesian Optimization: Generative Models as Acquisition Functions

ResearchDGX agent

arXiv:2510.25240v3 Announce Type: replace-cross Abstract: We present a general strategy for turning generative models into candidate solution samplers for batch Bayesian optimization (BO). The use of

Hyperspectral Image Land Cover Captioning Dataset for Vision Language Models

Model ReleasesDGX agent

arXiv:2505.12217v2 Announce Type: replace Abstract: We introduce HyperCap, the first large-scale hyperspectral captioning dataset designed to enhance model performance and effectiveness in remote sens

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models

Model ReleasesDGX agent

arXiv:2509.25826v3 Announce Type: replace Abstract: Inherent temporal heterogeneity, such as varying sampling densities and periodic structures, has posed substantial challenges in zero-shot generaliz

Mechanistic Interpretability of EEG Foundation Models via Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2605.13930v1 Announce Type: new Abstract: EEG foundation models achieve state-of-the-art clinical performance, yet the internal computations driving their predictions remain opaque: a barrier to

← Previous
1…8889909192…1009
Next →