AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
Model Releases

WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing

DGX agent

arXiv:2510.15221v2 Announce Type: replace Abstract: Affective computing has matured rapidly in laboratory settings, yet no prior dataset combines (i) months-to-years of duration, (ii) a naturalistic w

model-releasesarxiv-cs-ai
19 May 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

DGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

model-releasesarxiv-cs-ai
19 May 2026
Applications

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?

DGX agent

arXiv:2512.24497v3 Announce Type: replace Abstract: A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and env

applicationsarxiv-cs-ai
19 May 2026
Research

What is Holding Back Latent Visual Reasoning?

DGX agent

arXiv:2605.18445v1 Announce Type: cross Abstract: Humans can approach complex visual problems by mentally simulating intermediate visual steps, rather than reasoning through language alone. Inspired b

researcharxiv-cs-ai
19 May 2026
Research

What is the long-run distribution of stochastic gradient descent? A large deviations analysis

DGX agent

arXiv:2406.09241v3 Announce Type: replace-cross Abstract: In this paper, we examine the long-run distribution of stochastic gradient descent (SGD) in general, non-convex problems. Specifically, we see

researcharxiv-cs-lg
19 May 2026
Research

What Matters for Grocery Product Retrieval with Open Source Vision Language Models

DGX agent

arXiv:2605.18029v1 Announce Type: new Abstract: Multimodal product retrieval (MPR) underpins checkout-free retail and automated inventory systems, yet it demands fine-grained SKU discrimination that s

researcharxiv-cs-cv
19 May 2026
Research

When a Zero-Shooter Cheats: Improving Age Estimation via Activation Steering

DGX agent

arXiv:2605.17658v1 Announce Type: new Abstract: Different age-related regulations have been proposed to protect minors from harmful content and interactions online. Automated age estimation is central

researcharxiv-cs-lg
19 May 2026
Model Releases

When Accuracy Is Not Enough: Uncertainty Collapse between Noisy Label Learning and Out-of-Distribution Detection

DGX agent

arXiv:2605.17795v1 Announce Type: cross Abstract: Learning with noisy labels (LNL) is typically benchmarked by closed-set classification accuracy, yet deployment often requires classifiers to reject o

model-releasesarxiv-cs-cv
19 May 2026
Agents

When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning

DGX agent

arXiv:2605.16312v1 Announce Type: cross Abstract: We study adversarial action masking in self-play reinforcement learning: an attacker selectively removes legal actions from a victim's action set. Unl

agentsarxiv-cs-ai
19 May 2026
Model Releases

When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings

DGX agent

arXiv:2605.16288v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in clinical and care settings. This exploratory study investigates whether LLMs exhibit sycophantic

model-releasesarxiv-cs-cl
19 May 2026
Research

When Bits Break Recourse: Counterfactual-Faithful Quantization

DGX agent

arXiv:2605.17160v1 Announce Type: cross Abstract: Quantization can preserve predictive accuracy under low-bit deployment while silently breaking algorithmic recourse: an actionable change that flips a

researcharxiv-cs-ai
19 May 2026
Safety

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited

DGX agent

arXiv:2605.17017v1 Announce Type: cross Abstract: Behavior Foundation Models (BFMs) enable scalable imitation learning (IL) by pretraining task-agnostic representations that can be rapidly adapted to

safetyarxiv-cs-ai
19 May 2026
Research

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack

DGX agent

arXiv:2605.17288v1 Announce Type: cross Abstract: Large Language Model (LLM) cascade systems are designed to balance efficiency and performance by processing queries with lightweight models while sele

researcharxiv-cs-ai
19 May 2026
Research

When Fireflies Cluster; Enhancing Automatic Clustering via Centroid-Guided Firefly Optimization

DGX agent

arXiv:2605.18460v1 Announce Type: new Abstract: This work presents a novel variant of the Firefly Algorithm (FA) for data clustering, addressing limitations of traditional methods like K-Means that st

researcharxiv-cs-ai
19 May 2026
Safety

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search

DGX agent

arXiv:2605.16362v1 Announce Type: cross Abstract: Activation steering offers a lightweight way to control LLMs without retraining, but its effectiveness varies sharply across concepts. Prior work ofte

safetyarxiv-cs-ai
19 May 2026
Research

When Marginals Match but Structure Fails: Covariance Fidelity in Generative Models

DGX agent

arXiv:2603.17041v2 Announce Type: replace-cross Abstract: Generative models are increasingly deployed as substitutes for real data in downstream scientific workflows, yet standard evaluation criteria

researcharxiv-cs-ai
19 May 2026
Local Ai

When Molecular Similarity Works: Property Cliffs Reveal Hidden Errors

DGX agent

arXiv:2605.17265v1 Announce Type: new Abstract: Accurate prediction of molecular properties underpins drug discovery and material design, yet even state-of-the-art models remain vulnerable to localize

local-aiarxiv-cs-lg
19 May 2026
Model Releases

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

DGX agent

arXiv:2605.18580v1 Announce Type: new Abstract: Outcome-only evaluation can certify economically unsafe agents: a policy can hit a business KPI while violating deployable behavioral discipline. In hot

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents

DGX agent

arXiv:2601.17887v2 Announce Type: replace Abstract: Long-term memory enables large language model (LLM) agents to support personalized and sustained interactions. However, most work on personalized ag

model-releasesarxiv-cs-ai
19 May 2026
Tutorials

When TableQA Meets Noise: A Dual Denoising Framework for Complex Questions and Large-scale Tables

DGX agent

arXiv:2509.17680v2 Announce Type: replace Abstract: Table question answering (TableQA) is a fundamental task in natural language processing (NLP). The strong reasoning capabilities of large language m

tutorialsarxiv-cs-cl
19 May 2026
Safety

When Vision Speaks for Sound

DGX agent

arXiv:2605.16403v1 Announce Type: new Abstract: Despite rapid progress in video-capable MLLMs, we find that their apparent audio understanding in videos is often vision-driven: models rely on visual c

safetyarxiv-cs-cv
19 May 2026
Model Releases

Where Does Warm-Up Come From? Adaptive Scheduling for Norm-Constrained Optimizers

DGX agent

arXiv:2602.05813v2 Announce Type: replace Abstract: We study adaptive learning rate scheduling for norm-constrained optimizers (e.g., Muon and Lion). We introduce a generalized smoothness assumption u

model-releasesarxiv-cs-lg
19 May 2026
Safety

Where Pretraining writes and Alignment reads: the asymmetry of Transformer weight space

DGX agent

arXiv:2605.16600v1 Announce Type: cross Abstract: Cross-entropy pretraining and preference alignment update the same transformer weights, but leave geometrically distinct traces. We characterise this

safetyarxiv-cs-ai
19 May 2026
Safety

Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework

DGX agent

arXiv:2605.18150v1 Announce Type: new Abstract: Diffusion models (DMs) are widely used for text-to-image generation, but their strong generative capabilities also raise concerns about unsafe or undesi

safetyarxiv-cs-ai
19 May 2026
Safety

White-Box Sensitivity Auditing with Steering Vectors

DGX agent

arXiv:2601.16398v2 Announce Type: replace-cross Abstract: Algorithmic audits are essential tools for examining systems for properties required by regulators or desired by operators. Current audits of

safetyarxiv-cs-cl
19 May 2026
Model Releases

WhiteTesseract: Reframing the Interpretation of Cultural Heritage through XR and Conversational AI

DGX agent

arXiv:2605.16972v1 Announce Type: cross Abstract: Cultural heritage exhibitions often struggle to sustain attention and support reflective engagement. Physical exhibitions rely on fixed interpretive a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

DGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

model-releasesarxiv-cs-ai
19 May 2026
Research

Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road

DGX agent

arXiv:2605.17026v1 Announce Type: new Abstract: Recent progress in large language models has led to the emergence of reasoning models, which have shown strong performance on complex tasks through spec

researcharxiv-cs-lg
19 May 2026
Safety

Why Do Safety Guardrails Degrade Across Languages?

DGX agent

arXiv:2605.17173v1 Announce Type: cross Abstract: Large language models exhibit safety degradation in non-English languages. Standard evaluation relies on Jailbreak Success Rate (JSR), which confounds

safetyarxiv-cs-ai
19 May 2026
Research

Why Modeling Human Haptic Material Perception with AI Is Difficult

DGX agent

arXiv:2605.16602v1 Announce Type: cross Abstract: Touch plays a central role in how humans perceive and recognize materials through physical contact. Despite decades of research, the mechanisms by whi

researcharxiv-cs-ai
19 May 2026
Agents

Why We Look Where We Look: Emergent Human-like Fixations of a Foveated Visual Language Model Maximizing Scene Understanding

DGX agent

arXiv:2605.17823v1 Announce Type: cross Abstract: When humans view scenes without a specific task (free-viewing), they initially direct their eye movements toward the scene center and then fixate on p

agentsarxiv-cs-ai
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Research

WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points

DGX agent

arXiv:2605.17471v1 Announce Type: new Abstract: Quantization-aware training (QAT) is widely adopted to quantize language models by training full-precision weights using gradients from the quantized mo

researcharxiv-cs-lg
19 May 2026
Model Releases

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens

DGX agent

arXiv:2605.18115v1 Announce Type: new Abstract: Building a unified visual tokenizer is essential for bridging the gap between visual understanding and generation. Yet existing approaches struggle with

model-releasesarxiv-cs-cv
19 May 2026
Safety

World Model-Enabled Causal Digital Twins for Semantic Communications in Physical AI Systems

DGX agent

arXiv:2605.16547v1 Announce Type: new Abstract: Semantic communication has emerged as a promising paradigm for enabling goal-oriented networking. However, most existing semantic communication solution

safetyarxiv-cs-lg
19 May 2026
Model Releases

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

DGX agent

arXiv:2605.17912v1 Announce Type: cross Abstract: World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about envir

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

WOW-Seg: A Word-free Open World Segmentation Model

DGX agent

arXiv:2605.16903v1 Announce Type: new Abstract: Open world image segmentation aims to achieve precise segmentation and semantic understanding of targets within images by addressing the infinitely open

model-releasesarxiv-cs-cv
19 May 2026
Applications

XCTFormer: Leveraging Cross-Channel and Cross-Time Dependencies for Enhanced Time-Series Analysis

DGX agent

arXiv:2605.18534v1 Announce Type: new Abstract: Multivariate time-series analysis involves extracting informative representations from sequences of multiple interdependent variables, supporting tasks

applicationsarxiv-cs-lg
19 May 2026
Agents

Xiaomi EV World Model: A Joint World Model Integrating Reconstruction and Generation for Autonomous Driving

DGX agent

arXiv:2605.18137v1 Announce Type: new Abstract: This report presents a unified technical system addressing the two core capabilities of world models for autonomous driving: world representation and wo

agentsarxiv-cs-cv
19 May 2026
Local Ai

YawDD+: Frame-level Annotations for Accurate Yawn Prediction

DGX agent

arXiv:2512.11446v3 Announce Type: replace Abstract: Driver fatigue remains a leading cause of road accidents, responsible for 24% of crashes. While yawning serves as an early behavioral indicator of f

local-aiarxiv-cs-cv
19 May 2026
Model Releases

YOLO-NAS-Bench: A Surrogate Benchmark with Self-Evolving Predictors for YOLO Architecture Search

DGX agent

arXiv:2603.09405v2 Announce Type: replace Abstract: Neural Architecture Search (NAS) for object detection is severely bottlenecked by high evaluation cost, as fully training each candidate YOLO archit

model-releasesarxiv-cs-cv
19 May 2026
Research

You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations

DGX agent

arXiv:2511.06516v3 Announce Type: replace Abstract: Many LLM applications require only narrow capabilities, yet standard post-training quantization (PTQ) methods allocate precision without considering

researcharxiv-cs-cl
19 May 2026
Model Releases

Your SaaS Is an Insurance Product: A Modeling Framework

DGX agent

arXiv:2605.16699v1 Announce Type: new Abstract: Capped-usage SaaS products -- LLM subscriptions such as Claude Code and ChatGPT, cloud platforms such as Vercel and Cloudflare Workers, corporate benefi

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Zero-Shot Faithful Textual Explanations via Directional-Derivative Influence on Predictions

DGX agent

arXiv:2605.16877v1 Announce Type: new Abstract: Zero-shot textual explanations aim to make image classifiers more transparent by probing their internal representations, without relying on task-specifi

model-releasesarxiv-cs-cv
19 May 2026
Safety

Zero-Shot Textual Explanations via Translating Decision-Critical Features

DGX agent

arXiv:2512.07245v2 Announce Type: replace Abstract: Textual explanations make image classifier decisions transparent by describing the prediction rationale in natural language. Large vision-language m

safetyarxiv-cs-cv
19 May 2026
Safety

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse

DGX agent

arXiv:2509.23183v3 Announce Type: replace Abstract: Test-time entropy minimization helps adapt a model to novel environments and incentivize its reasoning capability, unleashing the model's potential

safetyarxiv-cs-lg
19 May 2026
Research

2Mamba2Furious: Linear in Complexity, Competitive in Accuracy

DGX agent

arXiv:2602.17363v3 Announce Type: replace Abstract: Linear attention transformers have become a strong alternative to softmax attention due to their efficiency. However, linear attention tends to be l

researcharxiv-cs-lg
18 May 2026
Model Releases

3D Segmentation Using Viewpoint-Dependent Spatial Relationships

DGX agent

arXiv:2605.15708v1 Announce Type: new Abstract: Recent advances in 3D datasets and multimodal models have greatly improved natural language 3D scene understanding. However, most 3D referring segmentat

model-releasesarxiv-cs-cv
18 May 2026
← Previous
1…856857858859860…1300
Next →