AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,572 results
28 Jul 2026

In-Context Learning as Implicit Policy Gradient

SafetyDGX agent

arXiv:2607.23153v1 Announce Type: cross Abstract: Recent work has shown that large language models (LLMs) can iteratively improve their outputs by incorporating generated samples and their correspondi

Inference-Time Consensus for Mitigating Hidden Behaviors from LLM Fine-Tuning

ResearchDGX agent

arXiv:2607.23394v1 Announce Type: new Abstract: Recent work shows that fine-tuning language models on even a small amount of poisoned data can install targeted misbehavior, and ostensibly benign data

Infrared Organization and Critical Cognitive Field Formation in Transformer Dynamics

Local AiDGX agent

arXiv:2607.10923v2 Announce Type: replace Abstract: Large language models exhibit remarkable emergent behaviors, yet the physical mechanism governing their collective dynamics remains poorly understoo

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Inverse Bayesian Inference for Extracting Lesion Dynamics from Longitudinal Spectral CT

Model ReleasesDGX agent

arXiv:2607.23078v1 Announce Type: new Abstract: Longitudinal medical imaging captures temporal evolution of lesions, yet extracting the underlying dynamical parameters governing this evolution remains

JPEG AIC2026: A large-scale dataset for fine-grained assessment of image coding

Model ReleasesDGX agent

arXiv:2607.22783v1 Announce Type: cross Abstract: Recent advances in conventional and learning-based image coding have increased the demand for benchmark datasets that support fine-grained assessment

LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories

Model ReleasesDGX agent

arXiv:2607.23704v1 Announce Type: cross Abstract: The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their reliability is constrained by the irre

Language-Routed RAG and Direct Option Scoring for Multilingual Financial QA: DS@GT at FinMMEval

Model ReleasesDGX agent

arXiv:2607.22841v1 Announce Type: cross Abstract: We present DS@GT's submission to FinMMEval 2026 Task 1, a multilingual financial exam question answering benchmark spanning English, Spanish, Greek, C

Latent-LoRA: Compact Latent-Space Adapters with Gradient-Free Routing for Continual Learning

ResearchDGX agent

arXiv:2607.23837v1 Announce Type: cross Abstract: Large language models generalize well to individual tasks but lack an inherent mechanism for learning them sequentially, leading to catastrophic forge

Learning-based Hierarchical Tracheal Anatomy Understanding from Sparse Surgical Demonstration Annotations for Ultrasound Robots

Local AiDGX agent

arXiv:2607.22789v1 Announce Type: cross Abstract: Tracheostomy requires precise localization of the tracheal incision site; however, conventional manual palpation is subjective and often unreliable, w

Learning to Access Computation: Accessibility Plasticity as a Principle of Adaptive Intelligence

Model ReleasesDGX agent

arXiv:2607.22748v1 Announce Type: cross Abstract: Modern neural networks primarily adapt through parameter modification within predefined computational structures. While recent methods introduce modul

LLM-Assisted Ontology Engineering and Construction of a French Legal Knowledge Graph

Model ReleasesDGX agent

arXiv:2607.24551v1 Announce Type: new Abstract: Maintenance regulations are complex legal texts that are difficult to exploit when addressing a specific case and challenging to integrate into operatio

Looking for Affect in Spontaneous Finnish Speech through Linguistic Interpretability

Model ReleasesDGX agent

arXiv:2607.24155v1 Announce Type: new Abstract: Existing research on affect in speech has shown how acoustic surface characteristics and content-related linguistic aspects of speech both relate to per

LoTA-N2N: Local Trace Adaptation for Zero-Shot Self-Supervised Image Denoising

Model ReleasesDGX agent

arXiv:2607.24135v1 Announce Type: new Abstract: Single-image self-supervised denoising replaces unavailable clean targets with surrogate targets constructed from noisy observations. Its effectiveness

MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning

Local AiDGX agent

arXiv:2607.24064v1 Announce Type: cross Abstract: Recent Vision-Language Models (VLMs) have achieved remarkable success in visual understanding, driven by the growing availability of high-quality imag

MATS: A novel multi-modality multi-task learning framework for 3D perception in autonomous driving

Model ReleasesDGX agent

arXiv:2607.24224v1 Announce Type: new Abstract: Multi-modality data from different sensors provides rich complementary information for 3D perception, becoming an essential component in reliable autono

MemChain: Learning Interpretable Memory Traces for Memory-Augmented LLM Agents

SafetyDGX agent

arXiv:2607.24097v1 Announce Type: new Abstract: Memory-augmented LLM agents typically answer queries by retrieving relevant memories and feeding them directly to an answer model. This retrieval-as-evi

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

Model ReleasesDGX agent

arXiv:2607.23811v1 Announce Type: cross Abstract: Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple's most powerful

MoE-Based Learned Inertial Odometry for Bicycle Localization

Model ReleasesDGX agent

arXiv:2510.17604v2 Announce Type: replace Abstract: GNSS suffers from multipath errors in urban canyons, making reliable bicycle localization difficult. Hand-crafted inertial alternatives, such as cyc

MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents

Model ReleasesDGX agent

arXiv:2607.23870v1 Announce Type: cross Abstract: Smart-city airspace is transforming Uncrewed Aerial Vehicles (UAVs) from passive sensing platforms into cyber-physical decision makers that must follo

Multi-Domain Physics-Based MDO of Multirotor UAVs: A Deterministic Framework for Discrete COTS Sizing

Model ReleasesDGX agent

arXiv:2607.22768v1 Announce Type: cross Abstract: Multirotor Unmanned Aerial Vehicle (UAV) design is governed by a tightly coupled system of non-linear equations spanning structural mechanics, electro

NEO: NeRF It Once, Edit It Many Times for Continuous Object Manipulation

Model ReleasesDGX agent

arXiv:2607.24538v1 Announce Type: new Abstract: In this paper, we present NEO, a unified framework providing language-guided NeRF editing for robotic manipulation. Our paper introduces (i) a language-

OmniMate: Open-Ended Real-Time Streaming Audio-Visual Generation for Interactive Avatars

ResearchDGX agent

arXiv:2607.23023v1 Announce Type: new Abstract: Recent advances in diffusion-based generative models have enabled real-time audio-driven avatar generation and unified audio-visual synthesis, providing

PathSelect: Sequential Token Selection for Whole Slide Pathology

TutorialsDGX agent

arXiv:2607.23631v1 Announce Type: new Abstract: Gigapixel Whole-Slide Images (WSIs) present a fundamental computational bottleneck for vision-language models (VLMs) due to extreme sequence lengths. Ex

PeopleSearchBench: A Multi-Dimensional Benchmark for Evaluating AI-Powered People Search Platforms

Model ReleasesDGX agent

arXiv:2603.27476v2 Announce Type: replace Abstract: AI-powered people search platforms are increasingly used in recruiting, sales prospecting, and professional networking, yet no widely accepted bench

Physics-Informed Neural Networks for Predicting Nitrous Oxide Flux

ResearchDGX agent

arXiv:2607.23880v1 Announce Type: cross Abstract: Nitrous oxide (N_2O) is the dominant ozone-depleting substance emitted in the 21st century, and the third largest contributor to anthropogenic greenho

QFedPolyp: A Communication- and Inference-Efficient Federated Learning Framework for Polyp Segmentation

Local AiDGX agent

arXiv:2607.22743v1 Announce Type: cross Abstract: Background and Objective: Automatic polyp segmentation supports computer-aided diagnosis and early colorectal cancer detec- tion. Centralized deep lea

Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline

HardwareDGX agent

arXiv:2607.22997v1 Announce Type: cross Abstract: Physical AI -- the integration of large vision-language-action (VLA) models with embodied agents that act in the real world -- has emerged as the next

Recurrent Autoregressive Diffusion: Global Memory Meets Local Attention

Local AiDGX agent

arXiv:2511.12940v2 Announce Type: replace Abstract: Recent advancements in video generation has shifted from bidirectional models for short videos to autoregressive ones for ultra long video generatio

Schema-Aware Localisation (SAL): Live Schema Grounding and Hallucination Validation for Oracle NL2SQL

Local AiDGX agent

arXiv:2607.22572v1 Announce Type: new Abstract: Large language models can generate fluent SQL from natural language, but on real enterprise Oracle databases they frequently fail at execution time: col

Shift-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment

SafetyDGX agent

arXiv:2501.19060v4 Announce Type: replace Abstract: Vision-language models (VLMs), such as CLIP, adapt effectively to downstream tasks through prompt tuning, but fine-tuning can misalign predictive co

SILICA: Repurposing Diffusion Priors for Joint Glass Segmentation and Depth Estimation

Model ReleasesDGX agent

arXiv:2607.24249v1 Announce Type: new Abstract: Standard depth sensors systematically fail on transparent surfaces, creating corrupted 3D maps and severe navigation hazards. While specialized hardware

SIREN: Towards End-to-End Extreme-Weather Early Warning with Experience-Grounded LLM Agents

Model ReleasesDGX agent

arXiv:2607.24588v1 Announce Type: new Abstract: Early warning of extreme weather is essential for mitigating the societal, economic, and environmental risks posed by hazardous weather events. However,

Small, Bias-Free, Blind and Convolutional Denoiser: A compact ConvNeXt U-Net for blind Gaussian color-image denoising

Model ReleasesDGX agent

arXiv:2607.22793v1 Announce Type: cross Abstract: We describe and evaluate BF-ConvUNeXt, a compact bias-free ConvNeXt U-Net for blind additive-white-Gaussian-noise color image denoising, combining fou

Snowflake debuts Cortex AI Gateway to govern and monitor enterprise AI agents

AgentsDGX agent

Snowflake Inc. today introduced Cortex AI Gateway, a centralized control layer that lets enterprises connect, govern and monitor artificial intelligence agents as they reach into models, tools, Model

Sources: Moonshot is seeking access to more Nvidia Blackwell chips to prepare for Kimi K4's development, after training K3 on Nvidia chips, including Blackwell (The Information)

Model ReleasesDGX agent

The Information: Sources: Moonshot is seeking access to more Nvidia Blackwell chips to prepare for Kimi K4's development, after training K3 on Nvidia chips, including Blackwell — Beijing-based startup

Speed Reading Tool Powered by Artificial Intelligence for Students with ADHD, Dyslexia, and Short Attention Span

Model ReleasesDGX agent

arXiv:2307.14544v2 Announce Type: replace-cross Abstract: This paper presents an artificial intelligence tool designed to assist students with dyslexia, ADHD, and short attention spans in processing t

STAIF: A Stage-wise Optimization for Complex Instruction Following

SafetyDGX agent

arXiv:2607.22649v1 Announce Type: new Abstract: Following complex instructions with multiple explicit constraints remains a fundamental challenge for large language models (LLMs). Existing alignment m

Teacher Knows It Best: Spontaneous Symmetry Breaking and Tipping Points in Networked Langevin Dynamics AI Sycophancy

ResearchDGX agent

arXiv:2607.24304v1 Announce Type: cross Abstract: We formulate a statistical physics framework to model a networked stochastic dynamical system exhibiting bistability, driven by additive noise and soc

TEmBed-T: A Multi-Dimensional Benchmark for Table-Level Embeddings

Model ReleasesDGX agent

arXiv:2607.24130v1 Announce Type: cross Abstract: Tabular data is the dominant structured-data modality, and learning table representations has become a core research direction. Table-level embeddings

Test-Time Adaptation via Dual Distillation for Videos Under Severe Distribution Shifts

SafetyDGX agent

arXiv:2607.24611v1 Announce Type: new Abstract: Deep learning models have achieved state-of-the-art performance in several computer vision tasks. However, they experience severe performance degradatio

The Entropic Bound for Transformers: Why Static Rank Fails and Attention-Native Rank Recovers

SafetyDGX agent

arXiv:2607.23050v1 Announce Type: new Abstract: Neural scaling laws describe how loss decreases as models, data, and compute grow, but they do not answer a prior question: for a fixed task, what is th

There's been a talk about how LLMs are only advancing in verifiable areas like math or coding, but that isn't what the data suggests. As mod…

ApplicationsDGX agent

There's been a talk about how LLMs are only advancing in verifiable areas like math or coding, but that isn't what the data suggests. As models have gotten better at that, they are also better at solv

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation

ResearchDGX agent

arXiv:2607.23493v1 Announce Type: cross Abstract: Automated analysis of multimodal content on social networks has become a critical task for understanding public sentiment and information diffusion in

Toward Optimal Adenovirus Detection Using YOLO26

ResearchDGX agent

arXiv:2607.17799v2 Announce Type: replace Abstract: This study systematically benchmarks different data augmentation setups across the baseline YOLO26 model size variants to determine the most effecti

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation

SafetyDGX agent

arXiv:2607.23181v1 Announce Type: new Abstract: Vision-and-Language Navigation in continuous environments (VLN-CE) requires an agent to ground language in egocentric observations and plan in unseen sc

Towards High-Level Semantic Intelligence

ResearchDGX agent

arXiv:2607.24082v1 Announce Type: new Abstract: Recent advances in AI have substantially expanded its cognitive and reasoning capabilities. From the perspective of semantic complexity, the development

Transfer Learning Architectures for Scalable Multi-Fidelity Bayesian Optimization

Model ReleasesDGX agent

arXiv:2607.23404v1 Announce Type: new Abstract: Self-driving laboratories increasingly rely on multi-fidelity Bayesian optimization (MFBO) to balance cheap, approximate evaluations against scarce, exp

UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents

Model ReleasesDGX agent

arXiv:2508.00288v5 Announce Type: replace-cross Abstract: Aerial navigation is a fundamental yet underexplored capability in embodied intelligence, enabling agents to operate in large-scale, unstructu

VIPER: Visual In-Context Physics Reasoning for Physically Plausible Video Generation

TutorialsDGX agent

arXiv:2607.23472v1 Announce Type: new Abstract: Modern video generation models can synthesize visually compelling and temporally coherent clips, yet controlling their physical behavior remains difficu

What CLIP Knows but Cannot Say: Recovering Negation from Frozen Intermediate Features

SafetyDGX agent

arXiv:2607.23271v1 Announce Type: cross Abstract: Contrastive vision-language models such as CLIP map semantically opposite phrases (e.g., 'a dog' vs. 'not a dog') to nearly identical embeddings, rend

What Softmax Throws Away: Mass-Aware Attention for Evidence Accumulation

ResearchDGX agent

arXiv:2607.22781v1 Announce Type: cross Abstract: High task performance does not show whether a model retains prediction-relevant structural information in its internal representation. Temporal graph

Wired for Overconfidence: A Mechanistic Perspective on Inflated Verbalized Confidence in LLMs

ResearchDGX agent

arXiv:2604.01457v3 Announce Type: replace Abstract: Large language models are often not just wrong, but confidently wrong: when they produce factually incorrect answers, they tend to verbalize overly

27 Jul 2026

Breaking the Data Barrier in Learning Symbolic Computation: A Case Study on Variable Ordering Suggestion for Cylindrical Algebraic Decomposition

Model ReleasesDGX agent

arXiv:2601.13731v2 Announce Type: replace-cross Abstract: Symbolic computation, powered by modern computer algebra systems, has important applications in mathematical reasoning through exact deep comp

Color Me Correctly: Bridging Perceptual Color Spaces and Text Embeddings for Improved Diffusion Generation

SafetyDGX agent

arXiv:2509.10058v2 Announce Type: replace Abstract: Accurate color alignment in text-to-image (T2I) generation is critical for applications such as fashion, product visualization, and interior design,

Complexity Bounds and Approaches to Learning Projected Gradient Descent Solver Iterates

ResearchDGX agent

arXiv:2607.22467v1 Announce Type: new Abstract: Data scarcity poses a fundamental challenge in training generative models to produce initial guesses for parametric optimization problems that are other

Constraint-Driven Synthesis of Hyper Petri Nets

SafetyDGX agent

arXiv:2607.22062v1 Announce Type: cross Abstract: This paper addresses the modeling and synthesis of constrained robotic system behaviors using Petri nets (PNs). It investigates how to construct model

FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs

Local AiDGX agent

arXiv:2607.21624v1 Announce Type: cross Abstract: Transformer-based models have enabled unprecedented capabilities across language, vision, and multimodal tasks. On-device fine-tuning of transformer m

Good move by @JensenHuang. The Nvidia letter is well written and worth reading. As we saw with the OpenAI-Hugging Face hack, we need open mo…

HardwareDGX agent

Good move by @JensenHuang. The Nvidia letter is well written and worth reading. As we saw with the OpenAI-Hugging Face hack, we need open models and harnesses for defense. Lets stop believing the PR t

I want to run Kimi K3 at home, so I’m trying to make 2.8T-scale experimentation cheaper

Local AiDGX agent

Hey r/LocalLLaMA, I’m a retired engineer with a background in distributed computing, currently running a 1-person startup. Like many people here, I’d love to experiment with 2T+ MoE models locally. Th

InteractComp: Evaluating Search Agents With Ambiguous Queries

Model ReleasesDGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

← Previous
1…448449450451452…1060
Next →