AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
1 Jul 2026

Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue

SafetyDGX agent

arXiv:2606.31719v1 Announce Type: cross Abstract: In collaborative dialogue, shared perception does not guarantee shared interpretation. Mutual understanding must be established through interaction. W

Surrogate-Gated Generation and Foundation-Model Embeddings for Bayesian Materials Design

Model ReleasesDGX agent

arXiv:2606.28578v1 Announce Type: cross Abstract: Closed-loop materials discovery iterates between proposing candidate structures and evaluating their properties, and property evaluation dominates the

Test-Time Verification for Text-to-SQL via Outcome Reward Models

SafetyDGX agent

arXiv:2606.30851v1 Announce Type: cross Abstract: Improving the reliability of large language models (LLMs) at inference time is a central challenge in structured reasoning tasks such as Text-to-SQL.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
30 Jun 2026

AD-DAE: Alzheimer's Disease Progression Modeling with Unpaired Longitudinal MRI using Diffusion Auto-Encoders

ResearchDGX agent

arXiv:2511.05934v2 Announce Type: replace Abstract: Generative modeling frameworks have emerged as an effective approach to capture high-dimensional image distributions from large datasets without req

Adaptive Block Diffusion: Resolving Training-Inference Mismatch in Diffusion Language Models

SafetyDGX agent

arXiv:2606.29275v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) are typically trained under fixed context structures, restricting denoising to predetermined token subsets. This create

Analyzing Defensive Misdirection Against Model-Guided Automated Attacks on Agentic AI Systems

SafetyDGX agent

arXiv:2606.20470v2 Announce Type: replace-cross Abstract: Agentic AI systems increasingly rely on language-model components to interpret instructions, process external data, invoke tools, and coordina

BrainJanus: A Unified Model for Understanding and Generation across Brain, Vision, and Language

SafetyDGX agent

arXiv:2606.30319v1 Announce Type: new Abstract: Modeling the bidirectional correspondence between external sensory stimuli and internal neural activity has emerged as a critical frontier in neuroscien

Expert-guided Clinical Text Augmentation via Query-Based Model Collaboration

SafetyDGX agent

arXiv:2509.21530v2 Announce Type: replace Abstract: Data augmentation is a widely used strategy to improve model robustness and generalization by enriching training datasets with synthetic examples. W

Guided Unconditional and Conditional Generative Models for Super-Resolution and Inference of Quasi-Geostrophic Turbulence

ResearchDGX agent

arXiv:2507.00719v3 Announce Type: replace-cross Abstract: Typically, numerical simulations of Earth systems are coarse, and Earth observations are sparse and gappy. We apply four generative diffusion

Illuminating Unified Multimodal Model for Free-form Interleaved Text-Image Generation

ResearchDGX agent

arXiv:2606.30054v1 Announce Type: new Abstract: The advancement of generative AI models capable of producing text and image marks a critical step forward in the realm of multimodal intelligence, parti

Latent Noise Mask for Reducing Visual Redundancy in Multimodal Large Language Models

ResearchDGX agent

arXiv:2606.30168v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) often fail in fine-grained visual reasoning, as question-relevant visual cues are diluted by dense and redundan

Lie Group Diffusion Models for Hardware-Aware Quantum Circuit Synthesis

Local AiDGX agent

arXiv:2606.29636v1 Announce Type: cross Abstract: An important task in quantum computing is unitary circuit synthesis compatible with physical hardware constraints. This problem has a natural hybrid s

MF-UAVPose6D: A Model-Free Monocular 6-DoF Pose Estimation Framework for Fixed-Wing UAVs

Local AiDGX agent

arXiv:2606.29697v1 Announce Type: new Abstract: For uncrewed aerial vehicles (UAVs), estimating six-degree-of-freedom (6-DoF) poses is essential for airspace situational awareness, target tracking, an

Multi-Block Diffusion Language Models

ResearchDGX agent

arXiv:2606.29215v1 Announce Type: cross Abstract: Block Diffusion Language Models (BD-LMs) improve diffusion-based text generation with KV caching and flexible-length generation. A natural next step i

Personalized Additive Modeling for Multi-level Federated Learning

ResearchDGX agent

arXiv:2405.16472v2 Announce Type: replace Abstract: Contemporary AI faces the challenge of balancing generality with user-specific personalization. In federated learning (FL), this challenge is amplif

Replica Symmetry Breaking and Algorithmic Thresholds in Empirical Risk Minimization under Multi-Index Model

TutorialsDGX agent

arXiv:2606.28573v1 Announce Type: new Abstract: Modern machine learning models are trained by optimizing high-dimensional non-convex empirical risk functions. Such cost functions can have a multitude

Rethinking Generative Reconstruction Attacks against Graph Neural Network Models

Model ReleasesDGX agent

arXiv:2606.29748v1 Announce Type: new Abstract: The application of graph data in numerous disciplines raises the need for gathering and analyzing huge volumes of data, some of which is private and sen

Skin-R1: Clinical Knowledge-Guided Dermatological Diagnosis Using Vision-Language Models

ResearchDGX agent

arXiv:2511.14900v2 Announce Type: replace-cross Abstract: Vision--language models (VLMs) have recently shown promise for assisting clinical reasoning in dermatological diagnosis. However, their trustw

Synthetic Interaction Data for Scalable Personalization in Large Language Models

AgentsDGX agent

arXiv:2602.12394v2 Announce Type: replace Abstract: Personalized prompting offers large opportunities for deploying large language models (LLMs) to diverse users, yet existing prompt optimization meth

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.29892v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become indispensable for pushing Vision-Language-Action Models (VLAs) beyond static imitation learning. However, exist

Uncertainty Estimation in Pathology Foundation Models via Deep Mutual Learning

ResearchDGX agent

arXiv:2606.30020v1 Announce Type: new Abstract: Pathology foundation models (PFMs) offer generalizable representations for whole-slide image (WSI) analysis, yet their clinical adoption remains limited

29 Jun 2026

Beyond MoCap: Scaling Motion Tokenizers with Synthetic Human Motion for Generative Modeling

ResearchDGX agent

arXiv:2606.27547v1 Announce Type: new Abstract: Human motion generation models are fundamentally constrained by the limited diversity of motion capture datasets, which predominantly contain common, re

Calibrating Biophysical Models for Grape Phenology Prediction via Multi-Task Learning

ApplicationsDGX agent

arXiv:2508.03898v2 Announce Type: replace-cross Abstract: Accurate prediction of grape phenology is essential for timely vineyard management decisions, such as scheduling irrigation and fertilization,

Continual Memorization of Factoids in Language Models

ResearchDGX agent

arXiv:2411.07175v3 Announce Type: replace Abstract: As new knowledge rapidly accumulates, language models (LMs) with pretrained knowledge quickly become obsolete. A common approach to updating LMs is

Everyone is concerned about Chinese models being banned The entire ecosystem of successful AI startups with real revenue are all catching pr…

HardwareDGX agent

Dylan Patel discusses concerns within the AI startup ecosystem about potential Chinese AI model bans and their impact on companies generating significant revenue. The post appears to address industry-

Fine-tuning a multimodal large language model for clinician-grade autism behavioral scoring from short home videos

Model ReleasesDGX agent

arXiv:2606.27484v1 Announce Type: new Abstract: Autism spectrum disorder (ASD) affects 1 in 31 US children, yet median age at diagnosis exceeds four years. Artificial intelligence pipelines that provi

Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?

Local AiDGX agent

arXiv:2501.16947v2 Announce Type: replace Abstract: The advances in Vision-Language models (VLMs) offer exciting opportunities for robotic applications involving image geo-localization - the problem o

Let Language Constrain Geometry: Vision-Language Models as Semantic and Spatial Critics for 3D Generation

SafetyDGX agent

arXiv:2511.14271v2 Announce Type: replace Abstract: Text-to-3D generation has advanced rapidly, yet state-of-the-art models, encompassing both optimization-based and feed-forward architectures, still

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior

Model ReleasesDGX agent

arXiv:2606.28182v1 Announce Type: cross Abstract: Embodied agents operating in decentralized and partially observable environments have attracted growing attention in recent years. However, existing l

SHIFT: Motion Alignment in Video Diffusion Models with Adversarial Hybrid Fine-Tuning

SafetyDGX agent

arXiv:2603.17426v2 Announce Type: replace Abstract: Image-conditioned video diffusion models achieve impressive visual realism but often suffer from weakened motion fidelity, e.g., reduced motion dyna

Single and Multi Truth Data Fusion using Large Language Models

Model ReleasesDGX agent

arXiv:2606.28062v1 Announce Type: cross Abstract: Data fusion, also known as truth discovery, is a data integration problem that aims to determine the correct value or set of values for each attribute

Towards Evaluation of Implicit Software World Models in Coding LLMs

ApplicationsDGX agent

arXiv:2606.27406v1 Announce Type: cross Abstract: Software engineering, whether performed by humans or by AI agents, requires reasoning about how software behaves. We call the internal model that supp

Vision-Default, Prior-Override: Causal Mechanisms of Perception-Knowledge Conflict in Vision-Language Models

ResearchDGX agent

arXiv:2606.28273v1 Announce Type: new Abstract: Vision-language models must reconcile visual evidence with memorized world knowledge when the two conflict. How they resolve this conflict shapes the re

xAI Grok audio models now available on Vercel AI Gateway

ToolsDGX agent

xAI's Grok audio models are now accessible through the Vercel AI Gateway, expanding the platform's capabilities for developers to integrate advanced audio processing into their applications. This inte

28 Jun 2026

It's quite rational to regulate frontier API models, especially to get more transparency for the government, without regulating open-source …

ResearchDGX agent

It's quite rational to regulate frontier API models, especially to get more transparency for the government, without regulating open-source AI. Here's why: 1. The most dangerous AI systems right now a

26 Jun 2026

A Systematic Survey of Semantic Role Labeling in the Era of Pretrained Language Models

ResearchDGX agent

arXiv:2502.08660v4 Announce Type: replace Abstract: Semantic role labeling (SRL) is a central natural language processing task for understanding predicate-argument structures within texts and enabling

BrepCoder: A Unified Multimodal Large Language Model for Multi-task B-rep Reasoning

AgentsDGX agent

arXiv:2602.22284v3 Announce Type: replace Abstract: Recent advancements in deep learning have actively addressed complex challenges within the Computer-Aided Design (CAD) domain.However, most existing

ConflictScore: Identifying and Measuring How Language Models Handle Conflicting Evidence

Model ReleasesDGX agent

arXiv:2606.26437v1 Announce Type: cross Abstract: Existing metrics for factuality and faithfulness evaluate whether an answer is supported or contradicted by its grounding documents, but they fail to

How Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple Mitigation

ApplicationsDGX agent

arXiv:2606.27275v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly critical to digital library workflows, yet their ability to process historical language remains poorly und

Hybrid privacy-aware semantic search: SVD-truncated document geometry and CKKS-encrypted query reranking under a restricted threat model

Model ReleasesDGX agent

arXiv:2606.26373v1 Announce Type: cross Abstract: Dense embeddings power semantic search and retrieval-augmented generation, but embedding-inversion attacks can reconstruct source text from a vector:

MedPruner: Training-Free Hierarchical Token Pruning for Efficient 3D Medical Image Understanding in Vision-Language Models

ResearchDGX agent

arXiv:2603.11625v2 Announce Type: replace-cross Abstract: While specialized Medical Vision-Language Models (VLMs) have achieved remarkable success in interpreting 2D and 3D medical modalities, their d

Modeling Local, Global, and Cross-Modal Context in Multimodal 3D MRI

TutorialsDGX agent

arXiv:2606.26894v1 Announce Type: new Abstract: Brain MRI poses a fundamental challenge for machine learning: models must learn from high-dimensional 3D data spanning multiple co-registered modalities

Revealing Mammographic Phenotypes in Deep Learning Breast Cancer Risk Models

ResearchDGX agent

arXiv:2606.26431v1 Announce Type: cross Abstract: Mammogram-based deep learning models have improved breast cancer risk prediction, but the learned imaging patterns remain underexplored. Existing inte

TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models

Model ReleasesDGX agent

arXiv:2503.15283v2 Announce Type: replace Abstract: Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation.

Use What You Know: Causal Foundation Models with Partial Graphs

ResearchDGX agent

arXiv:2602.14972v2 Announce Type: replace Abstract: Estimating causal quantities traditionally relies on bespoke estimators tailored to specific assumptions. Recently proposed Causal Foundation Models

25 Jun 2026

A Leakage-Aware Comparative Benchmark of Machine Learning, Deep Learning, and Transformer Models for Reliable Leukemia Detection

Model ReleasesDGX agent

arXiv:2606.24944v1 Announce Type: cross Abstract: Automated classification of acute lymphoblastic leukemia (ALL) from peripheral blood smear images has often reported near-perfect performance on the C

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs

Model ReleasesDGX agent

arXiv:2601.17037v2 Announce Type: replace Abstract: We investigate visual reasoning limitations of both multimodal large language models (MLLMs) and image generation models (IGMs) by creating a novel

ATMA: Length-Invariant Language Modeling via Polar Attention and Gated-Delta Compression Memory

Local AiDGX agent

arXiv:2606.25156v1 Announce Type: new Abstract: Modern large language models based on softmax scaled-dot-product attention are constrained by their training sequence length: as the key-value sequence

BFMTrack: Latent Sequence Optimization for Physics-Based Motion Tracking with Behavioral Foundation Models

SafetyDGX agent

arXiv:2606.25056v1 Announce Type: new Abstract: Behavioral Foundation Models (BFMs) offer a promising path toward universal physics-based character control by organizing a rich repertoire of physicall

BrainAgent: A Large Language Model-Driven Multi-Agent Framework for Autonomous Brain Signal Understanding

Model ReleasesDGX agent

arXiv:2606.25400v1 Announce Type: new Abstract: Brain-Computer Interfaces (BCIs) and brain signal understanding are pivotal for clinical health and next-generation interactions. Despite this significa

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

SafetyDGX agent

arXiv:2606.25473v1 Announce Type: new Abstract: Autoregressive video diffusion with causal diffusion transformers has emerged as a major paradigm for real-time streaming video generation and action-co

Closed-Loop Graph Algorithm Execution with Small Language Models: Step Accuracy and Rollout Reliability

AgentsDGX agent

arXiv:2606.24980v1 Announce Type: new Abstract: Small language models offer an efficient alternative to large-scale systems, but their ability to execute structured algorithms over multiple dependent

Congrats! Open source GLM model is really a game changer! Extremely fast, cheap, and high quality!

ToolsDGX agent

Fireworks AI announced the release of an open source GLM model that offers significant improvements in speed, cost efficiency, and output quality compared to existing alternatives. The post suggests t

Expresso-AI: Explainable Video-Based Deep Learning Models for Depression Diagnosis

ResearchDGX agent

arXiv:2606.25606v1 Announce Type: new Abstract: Given the widespread prevalence of depression and its consequential impact on individuals and society, it is crucial to obtain objective measures for ea

Failure Modes of Large Language Models on Research-Level Mathematics: A Taxonomy and an Empirical Characterisation

Model ReleasesDGX agent

arXiv:2606.24902v1 Announce Type: cross Abstract: The 'First Proof' benchmark [1] posed ten research-level mathematics questions to the strongest publicly available LLMs and found them consistently wr

Hypergraph Normal World Models for Logical Visual Anomaly Detection

Local AiDGX agent

arXiv:2606.25368v1 Announce Type: new Abstract: Visual anomaly detection is often deployed with only normal training images. Most one-class detectors map test patches or features to a normal reference

PRISM: Feed-Forward Single-Image 3D Reconstruction via Geometric Warp-Residual Modeling

Model ReleasesDGX agent

arXiv:2606.25430v1 Announce Type: new Abstract: Reconstructing 3D scenes from a single image is a fundamental challenge in computer vision, with broad applications in virtual reality, robotics, and co

Spam and Sentiment Detection in Arabic Tweets Using MARBERT Model

ResearchDGX agent

arXiv:2606.25495v1 Announce Type: new Abstract: Saudi Telecom Company (STC) is among the most popular companies in Saudi Arabia, with many customers. Yet, there is still a big room for improvement in

What happens when you shift more of the context layer into the model weights themselves? Engram is building a neolab focused on memory and c…

TutorialsDGX agent

What happens when you shift more of the context layer into the model weights themselves? Engram is building a neolab focused on memory and continual learning. Let’s go @dan_biderman @realJessyLin! Fun

24 Jun 2026

AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approach

Model ReleasesDGX agent

arXiv:2606.24655v1 Announce Type: cross Abstract: The explosive growth and complexity of product data within the dynamic Brazilian e-commerce landscape demand robust and specialized methods for struct

← Previous
1…152153154155156…1010
Next →