AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
Human
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
10 Jul 2026

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

Model ReleasesDGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

Unified Face Attack Detection via Fine-Grained Semantic Guidance

SafetyDGX agent

arXiv:2607.08156v1 Announce Type: new Abstract: The growing applications of facial recognition systems are accompanied by increasingly diverse security threats. Existing datasets lack detailed textual

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery

Model ReleasesDGX agent

arXiv:2607.08267v1 Announce Type: new Abstract: Unmanned aerial vehicles (UAVs) increasingly rely on visual grounding capabilities to localize task-relevant targets from diverse instructions in comple

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Unit-Independent Low-Rate Wrist GSR Processing for Stress Detection Using Phasic nSCR Features

ResearchDGX agent

arXiv:2607.08007v1 Announce Type: cross Abstract: Galvanic skin response (GSR) is widely used for stress detection, but wrist-based GSR remains challenging because its absolute amplitude can differ su

Unlocking Temporal Generalization in Hamiltonian Video Dynamics Models

ResearchDGX agent

arXiv:2607.07763v1 Announce Type: new Abstract: World models are typically trained to predict discrete-time physical dynamics with a fixed step size baked into the model weights, preventing prediction

Unpaired Joint Distribution Modeling via Multi-Scale Image Representations

ApplicationsDGX agent

arXiv:2607.08198v1 Announce Type: new Abstract: This paper studies the problem of learning a joint distribution from marginal observations, which is inherently ill-posed due to the ambiguity of feasib

Unveiling Public Opinion: A Study of Sentiment Analysis Using LSTM and Traditional Models

ResearchDGX agent

arXiv:2607.07772v1 Announce Type: new Abstract: In this age of social media, sites like Twitter have become meeting places for people to share their views and feelings on a wide range of issues and cu

Using AI-based Learning Assistants in Higher Education: A Large-Scale Descriptive Analysis

ApplicationsDGX agent

arXiv:2607.08748v1 Announce Type: new Abstract: In this study, we present a large-scale descriptive analysis of the use of an AI-based learning assistant (Syntea) in higher education. Based on objecti

UtterTune: LoRA-Based Target-Language Pronunciation Edit and Control in Multilingual Text-to-Speech

ResearchDGX agent

arXiv:2508.09767v3 Announce Type: replace-cross Abstract: We propose UtterTune, a lightweight method for adapting a multilingual text-to-speech (TTS) system built on a large language model (LLM). It i

Validating LLMs in social science: Epistemic threats and emerging norms

SafetyDGX agent

arXiv:2607.07915v1 Announce Type: cross Abstract: Large language models (LLMs) are reshaping social science methodology. Researchers increasingly prompt language models to generate quantitative measur

Validity of LLMs as data annotators: AMALIA on authority

Model ReleasesDGX agent

arXiv:2607.08731v1 Announce Type: cross Abstract: A national language model offers a linguistic community its own instrument for measuring what its citizens say and value. Portugal's AMALIA, a publicl

Vanilla SGD with Momentum Survives Heavy-Tailed Noise: Convergence Analysis without Gradient Clipping or Normalization

ResearchDGX agent

arXiv:2607.08104v1 Announce Type: new Abstract: Stochastic gradient descent (SGD) is a cornerstone of modern optimization. While its performance under heavy-tailed noise is often addressed through spe

Variational Phasor Circuits for Phase-Native Brain-Computer Interface Classification

Model ReleasesDGX agent

arXiv:2603.18078v2 Announce Type: replace Abstract: We present the Variational Phasor Circuit (VPC), a deterministic classical learning architecture on the continuous S^1 unit-circle manifold. Inspire

VectorizationLLM: Smart Vectorization Based AI Assistant

TutorialsDGX agent

arXiv:2607.07846v1 Announce Type: new Abstract: VectorizationLLM is a specialized Large Language Model based on Google open-weight LLMs. The model is designed to assist students to learn smart vectori

VEGAS: Human-Aligned Video Caption Evaluation via Gaze

ResearchDGX agent

arXiv:2607.08489v1 Announce Type: cross Abstract: Vision-language models excel at video captioning, yet typically generate descriptions that fail to capture individual viewers' attention. We propose V

Vision-Language Memory for Spatial Reasoning

ResearchDGX agent

arXiv:2511.20644v2 Announce Type: replace Abstract: Spatial reasoning is a critical capability for intelligent robots, yet current vision-language models (VLMs) still fall short of human-level perform

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval

ResearchDGX agent

arXiv:2607.08541v1 Announce Type: cross Abstract: Open-vocabulary object detection and segmentation aim to recognize arbitrary objects beyond predefined categories. Although recent vision-language and

VSRo-200: A Romanian Visual Speech Recognition Dataset for Studying Supervision and Multimodal Robustness

Model ReleasesDGX agent

arXiv:2607.08112v1 Announce Type: new Abstract: We introduce VSRo-200, the first large-scale dataset for visual speech recognition (lip reading) in Romanian, comprising 200 hours of real-world podcast

WaspMOT: A Benchmark for Long-Term Multi-Object Tracking of Trichogramma Wasps

Model ReleasesDGX agent

arXiv:2607.08729v1 Announce Type: new Abstract: Multi-object tracking (MOT) has achieved strong performance on benchmarks dominated by short video sequences. However, such datasets do not adequately e

Wat3R: Underwater 3D Geometry Learning without Annotations

ResearchDGX agent

arXiv:2607.08772v1 Announce Type: new Abstract: Estimating 3D geometry in underwater environments presents unique challenges due to light attenuation, scattering, and the absence of large-scale, high-

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search

Local AiDGX agent

arXiv:2607.08662v1 Announce Type: cross Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-

Whareformer: Learning to Track What is Where in Long Egocentric Videos

ResearchDGX agent

arXiv:2607.08537v1 Announce Type: new Abstract: The recently established 'Out of Sight, Not out of Mind' (OSNOM) task for egocentric videos focuses on tracking objects that are moved by the camera wea

What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness

ResearchDGX agent

arXiv:2607.08046v1 Announce Type: cross Abstract: Large language models fine-tuned for forecasting can be accurate yet poorly calibrated, and their chain-of-thought (CoT) reasoning may not faithfully

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents

Model ReleasesDGX agent

arXiv:2607.08032v1 Announce Type: new Abstract: Large language models, and the agents built on them, spend an ever-growing share of their compute and memory on remembering: caching attention keys and

When Debiasing Backfires: Counterintuitive Side Effects of Preprocessing-Based Stereotype Mitigation

ResearchDGX agent

arXiv:2607.07937v1 Announce Type: new Abstract: Preprocessing-based methods for stereotype mitigation, such as pre-/post-training on debiased corpora, are widely used in NLP. While these approaches re

When Does Continual Learning Require Learning

AgentsDGX agent

arXiv:2607.07847v1 Announce Type: new Abstract: As large language models (LLMs) become increasingly capable, the next question is how can we enable models to continually learn? Today, the field largel

When Implausible Tokens Get Reinforced: Tail-Aware Credit Calibration for LLM Reinforcement Learning

ResearchDGX agent

arXiv:2607.07976v1 Announce Type: cross Abstract: Reinforcement learning (RL) has achieved remarkable success in enhancing the reasoning capabilities of large language models (LLMs). However, widely u

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

Model ReleasesDGX agent

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al.

When Structured Sparse Autoencoders Learn Consistent Concepts Across Modalities

SafetyDGX agent

arXiv:2607.08605v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a promising technique for mechanistic interpretability by learning a set of sparse latent features in large

When Synthetic Speech Is All You Have: Better Call GRPO

SafetyDGX agent

arXiv:2607.08409v1 Announce Type: cross Abstract: LLM-based ASR adapted to regulated domains such as banking is bottlenecked by privacy: real speech is costly and legally constrained to collect, makin

When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability

Model ReleasesDGX agent

arXiv:2607.08535v1 Announce Type: cross Abstract: An LLM-as-judge score can move even when the candidate responses stay fixed, simply because the evaluator has changed. We treat this evaluator-replace

When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models

Model ReleasesDGX agent

arXiv:2607.08059v1 Announce Type: cross Abstract: Uncertainty quantification for visual language models (VLMs) conventionally targets the answer token distribution. We provide the first three-family e

Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-STPA

Model ReleasesDGX agent

arXiv:2607.08054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly trusted to draft the artifacts of safety analysis such as, losses, hazards, Unsafe Control Actions (UCAs

Who Broke the System? Failure Localization in LLM-Based Multi-Agent Systems

Local AiDGX agent

arXiv:2607.07989v1 Announce Type: cross Abstract: Large language model (LLM) based multi-agent systems enable complex problem solving through coordinated reasoning and action, but their distributed st

Who Gets Missed in the Tail? Thresholded Subgroup Underdiagnosis in Long-Tailed Chest X-ray Classification

SafetyDGX agent

arXiv:2607.07717v1 Announce Type: cross Abstract: In chest X-ray (CXR) classification, acceptable ranking performance can still leave rare-positive patients below threshold, especially within subgroup

Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows

SafetyDGX agent

arXiv:2607.08740v1 Announce Type: new Abstract: Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Exist

Workload-Preserving Differentially Private Synthetic Data for Causal Inference via Maximum-Entropy Calibration

Model ReleasesDGX agent

arXiv:2607.08122v1 Announce Type: new Abstract: Workload-based differentially private (DP) synthetic data methods privately measure aggregate queries and post-process the noisy answers into synthetic

Write-Protected Discrete Bottlenecks for Language-Grounded World Models: A Structural Limitation and Sufficient Fix

TutorialsDGX agent

arXiv:2607.08312v1 Announce Type: new Abstract: How should language interface with a world model's discrete symbol system? The dominant paradigm -- end-to-end injection of LLM/VLM features into robot

X-ACTA: eXtended Analytic Center Tension distribution Algorithm for fixed and mobile cable-driven-parallel-robot

ResearchDGX agent

arXiv:2607.08265v1 Announce Type: new Abstract: Steering Cable-Driven Parallel Robots (CDPRs) beyond their Wrench-Feasible Workspace (WFW) augments their capabilities in challenging scenarios such as

XALPHA: A Memory-Driven AI Quant Researcher for Hypothesis-to-Code Alpha Discovery

SafetyDGX agent

arXiv:2607.08332v1 Announce Type: new Abstract: Financial markets are noisy, non-stationary, and high-dimensional, making it difficult to discover predictive and robust trading signals. Alpha discover

XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision

SafetyDGX agent

arXiv:2601.21688v2 Announce Type: replace-cross Abstract: Disentangled representation learning aims to map independent factors of variation to independent representation components. On one hand, purel

XOV-Action: Towards Generalizable Open-Vocabulary Action Recognition

Model ReleasesDGX agent

arXiv:2403.01560v3 Announce Type: replace Abstract: Inspired by the impressive success of image-text foundation models, recent works have proposed to adapt these foundation models to video data, leadi

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device

ResearchDGX agent

arXiv:2607.08771v1 Announce Type: new Abstract: Monocular depth estimation has seen remarkable progress through foundation models achieving robust zero-shot generalization, yet their computational dem

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning

SafetyDGX agent

arXiv:2601.02918v3 Announce Type: replace Abstract: Image Quality Assessment (IQA) is a long-standing problem in computer vision. Previous methods typically focus on predicting numerical scores withou

9 Jul 2026

A Closed-Loop Multi-Agent Framework for Robust Multi-Robot Manipulation

AgentsDGX agent

arXiv:2607.06990v1 Announce Type: new Abstract: Multi-robot systems provide the parallelism and redundancy necessary for long-horizon tasks, while Large Language Models (LLMs) offer the reasoning capa

A Continual Learning Framework for Adaptive Control of Modular Soft Robots

TutorialsDGX agent

arXiv:2607.06740v1 Announce Type: cross Abstract: Soft robots have attracted significant attention in applications such as medical intervention, rehabilitation, and robotic manipulation due to their i

A Distributionally Robust Optimisation Approach to Fair Credit Scoring

SafetyDGX agent

arXiv:2402.01811v2 Announce Type: replace Abstract: Credit scoring has been catalogued by the European Commission and the Executive Office of the US President as a high-risk classification task, in li

A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong

Model ReleasesDGX agent

arXiv:2607.06854v1 Announce Type: cross Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade,

A Good Initialization is All You Need for Faithful Visual Attribution

ResearchDGX agent

arXiv:2607.06726v1 Announce Type: new Abstract: Faithful visual attribution identifies which image regions support a model prediction. Search-based perturbation methods lead the insertion--deletion fa

A knowledge-augmented dataset of high-risk driving scenarios with LLM annotations for autonomous driving

Model ReleasesDGX agent

arXiv:2607.07103v1 Announce Type: new Abstract: Safe autonomous driving requires both rapid responses to common high-risk events and deeper reasoning over rare, extreme long-tail scenarios in traffic

A Multi-Analyst LLM Pipeline for Auditable Rule Discovery Across 68 Public Physiological Corpora

ResearchDGX agent

arXiv:2607.06802v1 Announce Type: cross Abstract: Open physiological corpora are heterogeneous: they use different sensors, labels, sampling rates, recording settings, and clinical endpoints. They can

A Quiet Failure in Calibrated Virtual Screening: Marginal Conformal Prediction Under-Covers the Minority Class, and a Class-Conditional Fix Recovers It

Local AiDGX agent

arXiv:2607.06605v1 Announce Type: new Abstract: Conformal prediction is being adopted in drug discovery to put an honest number on model reliability: pick an error rate alpha, and the method returns p

A Study of Commonsense Reasoning over Visual Object Properties

Model ReleasesDGX agent

arXiv:2508.10956v3 Announce Type: replace-cross Abstract: Inspired by human categorization, visual reasoning about object properties, such as physical attributes and functions, involves identifying an

A Theory of Contrastive Learning with Natural Images

TutorialsDGX agent

arXiv:2607.07470v1 Announce Type: new Abstract: Why does contrastive learning with simple images and augmentations yield useful representations for downstream tasks? We address this question by analyt

A Unified Detection Framework for AI-Related Content and Artifacts

ResearchDGX agent

arXiv:2607.07527v1 Announce Type: cross Abstract: Artificial intelligence (AI) is a double-edged sword: while it has achieved remarkable success across a wide range of domains, its deployment also cal

A Word-Level Digital Reader of the Prasthanatrayi with Sankara's Bhasya: Corpus, Method, and an Open, Offline Reading Aid for the Advaita Vedanta Canon

ResearchDGX agent

arXiv:2607.07282v1 Announce Type: new Abstract: The Prasthanatrayi -- the ten principal Upanisads, the Brahmasutra, and the Bhagavadgita, with Sankara's commentaries (bhasya) -- is the foundational co

AA-ViT: Anatomically Aware Vision Transformer with Structural and Frequency Guidance for Contrast Enhanced Brain MRI Synthesis

Local AiDGX agent

arXiv:2607.07553v1 Announce Type: new Abstract: Accurate tumour localization and diagnosis is a critical component of clinical care for brain cancers. Magnetic Resonance Imaging (MRI) is the most comm

Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning

ResearchDGX agent

arXiv:2607.07708v1 Announce Type: cross Abstract: Structure-property relationships are foundational to biology, chemistry and materials science, where function, reactivity and physical response emerge

Ace! Motion Planning of Professional-Level Table Tennis Serves with a Robot Arm

Model ReleasesDGX agent

arXiv:2607.06989v1 Announce Type: new Abstract: Table tennis, a dynamic, compact, and popular sport, has received significant attention as a robotics benchmark over the last decades. Most of the resea

← Previous
1…214215216217218…1005
Next →