AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,661 results
24 Jul 2026

Token-Level Entropy Reveals Demographic Disparities in Large Language Models

SafetyDGX agent

arXiv:2501.19337v5 Announce Type: replace Abstract: A name alone measurably reshapes a language model's next-token distribution before a single token is sampled. We measure full-vocabulary Shannon ent

TopoGuard: Graph Theory Based Defenses Against Split-Knowledge Attacks on RAG

ApplicationsDGX agent

arXiv:2607.20437v1 Announce Type: new Abstract: Production Retrieval Augmented Generation (RAG) systems rely on aggregating multiple external documents to answer complex queries. However, the retrieve

TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics

Model ReleasesDGX agent

arXiv:2602.19313v2 Announce Type: replace-cross Abstract: General-purpose robot learning requires dense, instruction-conditioned feedback that can distinguish meaningful task progress from stalled, fa

Content type
AllBlogX PostPaperYouTubeRedditGitHub

torchsom: The Reference PyTorch Library for Self-Organizing Maps

Model ReleasesDGX agent

arXiv:2510.11147v2 Announce Type: replace-cross Abstract: This paper introduces torchsom, an open-source Python library that provides a reference implementation of the Self-Organizing Map (SOM) in PyT

TOUR: A Trajectory-Level Unlearning Benchmark for Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.21111v1 Announce Type: cross Abstract: Offline Reinforcement Learning (RL) agents are trained on fixed behavioral trajectories, which makes trajectory-level deletion important when selected

Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry

Local AiDGX agent

arXiv:2607.21495v1 Announce Type: new Abstract: AI agents are increasingly created inside organizations by non-engineering users through low-code, no-code, and conversational development environments.

Toward cryptographically verifiable authorization for autonomous AI agents: A security hypothesis, preliminary formal model, and proof-of-concept implementation

SafetyDGX agent

arXiv:2607.21325v1 Announce Type: cross Abstract: Autonomous AI agents increasingly execute actions, invoke tools, and operate on protected resources with limited human oversight. Existing authenticat

Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.21496v1 Announce Type: cross Abstract: Cognitive impairment (CI) is a growing public health concern. Early and accurate diagnosis is critical for enabling timely intervention and improving

Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry

Model ReleasesDGX agent

arXiv:2607.20778v1 Announce Type: new Abstract: Weather forecasting foundation models (FMs) are increasingly fine-tuned to predict air quality, offering fast global pollution forecasts at lower comput

Towards a Certifying Grounder

ResearchDGX agent

arXiv:2607.21199v1 Announce Type: cross Abstract: Grounding, the translation of high-level theories into equivalent quantifier-free formulas, is a crucial step in declarative solving, yet it has so fa

Towards an Automated Test of LLM Security Knowledge

Model ReleasesDGX agent

arXiv:2607.18496v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for a range of software, hardware and human-centered security tasks. Consequently, LLM perf

Towards Capability-Aware Traversability Navigation for Unstructured Environments

ResearchDGX agent

arXiv:2607.20679v1 Announce Type: new Abstract: Estimating traversability in unstructured environments requires conditioning on robot embodiment, as the same terrain can be traversable for one platfor

Towards Faithful Graph Explanations with Synergistic Edge Effects via Granular Balls

Model ReleasesDGX agent

arXiv:2607.21381v1 Announce Type: new Abstract: Instance-level explanations aim to reveal the rationale behind a model's decisions for a specific graph. Previous methods explain graph neural networks

Towards Privacy-Preserving Federated Prompt Tuning under Data Heterogeneity: A Subspace-Decomposed Expert Approach

Local AiDGX agent

arXiv:2607.21417v1 Announce Type: new Abstract: Federated prompt tuning (FPT) enables collaborative adaptation of vision--language models (VLMs) using lightweight prompts. Existing methods often addre

Towards Robust Iris Recognition Through Occlusion Identification and Conditional Diffusion-Based Reconstruction

ResearchDGX agent

arXiv:2607.21545v1 Announce Type: new Abstract: Iris recognition is a reliable biometric approach that identifies individuals using the distinctive and stable texture of the iris. However, recognition

Traceable Scholarship: Page Anchors and Ariadne's Thread for Humanistic Inquiry in the Age of Generative AI

AgentsDGX agent

arXiv:2607.20916v1 Announce Type: new Abstract: Generative AI lets large language models produce scholarly-looking text within seconds, yet fluency does not equal valid explanation. The deepest risk i

Tractable Hierarchical Control of Autoregressive Language Models

ResearchDGX agent

arXiv:2607.20483v1 Announce Type: new Abstract: Constraining the generation of autoregressive large language models (LLMs) is an important component of integrating language models into formal systems.

Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers

HardwareDGX agent

arXiv:2512.16615v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) set the state of the art in visual generation, yet their quadratic self-attention cost fundamentally limits scaling to

Training Large Language Models for Self-Explanation Faithfulness

Model ReleasesDGX agent

arXiv:2607.21090v1 Announce Type: cross Abstract: We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated r

TransBiolab: A Real-World Multi-View Dataset of Cluttered Transparent Biomedical Objects

Model ReleasesDGX agent

arXiv:2607.21071v1 Announce Type: new Abstract: Autonomous biomedical laboratories increasingly rely on visual perception to recognize, localize, and manipulate transparent plasticware, yet high-quali

Transformer-Assisted LLM-Based Source Code Summarisation: to Enable More Secure Software Development

ResearchDGX agent

arXiv:2607.20933v1 Announce Type: cross Abstract: Neural Source Code Summarisation (NSCS) aims to generate natural language summaries of source code to improve developers' and maintainers' understandi

Transformer-based Diffusion models for Hydrological Time Series Probabilistic Imputation and Forecasting

ResearchDGX agent

arXiv:2607.21200v1 Announce Type: cross Abstract: The modeling of hydrometeorological time series with limited observations is a key challenge in the monitoring of hydro-systems and water resources, a

Transition-Related Potentials as Markers of Narrative Comprehension in Continuous EEG

ResearchDGX agent

arXiv:2607.20720v1 Announce Type: cross Abstract: Harnessing the potential of electroencephalography (EEG) for brain research is fundamentally limited by intrinsic noise and the diffuse projection of

True story: About 10 years ago there was a long article (NYT maybe?) about the end of trucking, estimating that automating truck driving wil…

TutorialsDGX agent

True story: About 10 years ago there was a long article (NYT maybe?) about the end of trucking, estimating that automating truck driving will wipe out something like 1%-2% of the GPD because a surpris

TwistedMerge: Certified Higher-Order Diagnostics and Abstention for Model Merging

SafetyDGX agent

arXiv:2607.20887v1 Announce Type: cross Abstract: Model merging combines independently trained or fine-tuned models, but pairwise alignability does not imply globally consistent alignment. We formulat

U-CFR: Uncertainty-Guided Cascade Forward Refinement for Interactive Segmentation

Model ReleasesDGX agent

arXiv:2607.20705v1 Announce Type: cross Abstract: Interactive image segmentation is critical for efficient image annotation; however, existing methods often require many corrective clicks or rely on p

Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement

ResearchDGX agent

arXiv:2607.20529v1 Announce Type: cross Abstract: Large Language Model (LLM) ensembles are increasingly used to improve reliability by combining predictions from multiple LLMs. However, existing aggre

UnDA: Unpaired Domain Alignment for Cross-Modal Knowledge Transfer in Medical Imaging

SafetyDGX agent

arXiv:2607.21546v1 Announce Type: new Abstract: Multimodal based approaches often outperform single modality approaches in downstream tasks as the different modalities provide complementary informatio

Understanding Critical Thinking in Generative Artificial Intelligence Use: Development, Validation, and Correlates of the Critical Thinking in AI Use Scale

SafetyDGX agent

arXiv:2512.12413v2 Announce Type: replace Abstract: Generative AI tools are increasingly embedded in everyday work and learning, yet their fluency, opacity, and propensity to hallucinate mean that use

Unified Video Dense Prediction from Disjoint Data

ResearchDGX agent

arXiv:2607.21592v1 Announce Type: new Abstract: Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragment

Unlearning Under Imbalance: Benchmarking Fairness in Multimodal LLM Unlearning

Model ReleasesDGX agent

arXiv:2607.21300v1 Announce Type: cross Abstract: Machine unlearning has emerged as a tool for removing personal data from trained models to comply with recent AI regulations. To evaluate unlearning e

Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Ghana

ResearchDGX agent

arXiv:2607.21559v1 Announce Type: new Abstract: A consensus anomaly detection framework was applied to monthly malaria surveillance data from Ghana (2014-2023) to identify atypical transmission patter

Unsupervised Metal Artifact Reduction in Dental CBCT using Fine-tuned Cycle-Consistent Adversarial Networks

ResearchDGX agent

arXiv:2607.20977v1 Announce Type: new Abstract: Metal artifacts generated by dental implants significantly degrade cone-beam computed tomography (CBCT) volumes, obscuring critical anatomical structure

URF: A Unified Robot Control-Policy Framework for Stable Contact Aware Manipulation

SafetyDGX agent

arXiv:2607.20912v1 Announce Type: new Abstract: Learning-based manipulation policies usually predict robot actions from sensory observations and leave their execution to a separate low-level controlle

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure

SafetyDGX agent

arXiv:2607.21151v1 Announce Type: new Abstract: As Video Large Language Models are increasingly deployed in real-world applications, ensuring their safety alignment has become critical. Counterintuiti

Verifier-First Evaluation of Agentic LLMs for Infrastructure-as-Code Generation

Model ReleasesDGX agent

arXiv:2607.20478v1 Announce Type: cross Abstract: Infrastructure-as-Code (IaC) generation from natural language requires satisfying provider schemas, dependency planning, and organizational policy con

VeriSimpl: Robust Optimization Modeling from Natural Language using Simplification-based Verification

ResearchDGX agent

arXiv:2607.20474v1 Announce Type: new Abstract: Natural language interfaces can greatly benefit the accessibility and usability of optimization modeling, and recent advances in large language models (

VibeVoice-ASR-BitNet Technical Report

Local AiDGX agent

arXiv:2607.21075v1 Announce Type: cross Abstract: We present VibeVoice-ASR-BitNet, a compressed variant of VibeVoice-ASR optimized for real-time inference on edge CPUs. We apply heterogeneous quantiza

Vision-Language-Policy Model for Dynamic Robot Task Planning

SafetyDGX agent

arXiv:2512.19178v2 Announce Type: replace-cross Abstract: Bridging the gap between natural language commands and autonomous execution in unstructured environments remains an open challenge for robotic

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?

Model ReleasesDGX agent

arXiv:2607.20868v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental ab

Visual Contrastive Self-Distillation

Model ReleasesDGX agent

arXiv:2607.21556v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) is promising as it removes the external teacher required by on-policy distillation (OPD), yet it still needs asymme

VoLN: Vision-Only Long-Horizon Navigation---Paradigm, Benchmark, and Method

Model ReleasesDGX agent

arXiv:2607.21400v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) enables embodied agents to follow natural-language instructions. However, route-level instructions commonly encod

VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory

Model ReleasesDGX agent

arXiv:2603.04910v2 Announce Type: replace-cross Abstract: Imitation learning from human demonstrations has achieved significant success in robotic control, yet most visuomotor policies still condition

Was great to talk about this timely and alarming news, thanks for having me on!

SafetyDGX agent

Was great to talk about this timely and alarming news, thanks for having me on! . @NPCollapse, executive director at ControlAI, joins 'On Balance' to discuss the dangers of artificial intelligence aft

WAT3R: Feedforward Underwater 3D Reconstruction

ResearchDGX agent

arXiv:2607.21023v1 Announce Type: new Abstract: Reliable feedforward underwater 3D reconstruction remains challenging due to severe light attenuation and backscattering, which degrade visual quality a

WaveformQA: Benchmarking LLM Temporal Reasoning on Digital Waveforms

Model ReleasesDGX agent

arXiv:2607.20638v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation and reasoning, yet their ability to perform temporal reasoning ove

We are pleased to announce that @GaryMarcus, Professor Emeritus at New York University, will speak at the 19th Annual AGI Conference (July 2…

SafetyDGX agent

We are pleased to announce that @GaryMarcus, Professor Emeritus at New York University, will speak at the 19th Annual AGI Conference (July 27–30, 2026). Professor Marcus is a leading voice in artifici

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few point…

Model ReleasesDGX agent

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing

WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance

Model ReleasesDGX agent

arXiv:2511.12997v2 Announce Type: replace Abstract: Multimodal LLM-powered agents have recently demonstrated impressive capabilities in web navigation, enabling agents to complete complex browsing tas

Webly Supervised Multi-Label Recognition: Evaluation Benchmark and Dual-Branch Multi-Label Contrastive Learning

Model ReleasesDGX agent

arXiv:2607.20874v1 Announce Type: new Abstract: Training deep learning models with freely available web images can reduce their dependence on costly manual annotations. Although webly supervised learn

Weight-norm Criticality: A Mechanism for Loss Spikes Induced by the Normalization and Weight Decay

Model ReleasesDGX agent

arXiv:2607.21005v1 Announce Type: new Abstract: Most explanations of training instability focus on learning-rate criticality, typically characterized by the Edge of Stability, beyond which optimizatio

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models.

SafetyDGX agent

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models. For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will

What I learned using Ollama on a real Paperless archive: model choice was not the main problem

Local AiDGX agent

I maintain Tagvico, an open-source companion for Paperless-ngx. I added Ollama because document text is exactly the kind of data many people do not want to send to a hosted model. The surprising failu

What if AI had access to classified files it was never allowed to quote but can make image?

IndustryDGX agent

What if an AI had seen fragments of classified material it could never describe directly? No files. No report names. No official explanations. Just images. That was the concept behind this series. I a

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

Model ReleasesDGX agent

arXiv:2607.20425v1 Announce Type: new Abstract: What makes writing 'good' remains a persistent question in literary studies and computational linguistics. We present a two-study investigation of how r

What Matters for Simulation to Online Reinforcement Learning on Real Robots

ApplicationsDGX agent

arXiv:2602.20220v2 Announce Type: replace-cross Abstract: We investigate what specific design choices enable successful online reinforcement learning (RL) on physical robots. Across 100 real-world tra

What, Where, and How: Disentangling the Roles of Task, Language, and Model in Code Model Representations

Model ReleasesDGX agent

arXiv:2607.21491v1 Announce Type: new Abstract: Do independently trained language models come to represent the same thing in the same way? We answer for code, extending a recently introduced concept-c

What's the last model trained on human-data only?

Local AiDGX agent

From my understanding, most current LLMs are trained on trillions and trillions of tokens of mostly AI-generated data. Are there any recent models that are trained purely (or as close as possible) on

When Are Reasoning-Based Guardrails Not Efficient? ResponseGuard: A Fast Vision-Language Guard for Real-Time Moderation

Model ReleasesDGX agent

arXiv:2607.21401v1 Announce Type: cross Abstract: A vision-language AI assistant returns its answer as a stream of generated tokens. Therefore, a safety guard that watches that answer has to keep up w

When Does Recurrence Become an Algorithm? Convergence Selection in Weight-Tied Looped Transformers

Model ReleasesDGX agent

arXiv:2607.20594v1 Announce Type: cross Abstract: When does a weight-tied looped transformer -- one block applied T times -- implement an actual algorithm? We answer with four findings from controlled

← Previous
1…215216217218219…1412
Next →