AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Safety

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty

DGX agent

arXiv:2605.27288v1 Announce Type: cross Abstract: Large language models (LLMs) are known to abandon their initial stance to conform to user pushback. While prior research largely attributes this behav

safetyarxiv-cs-ai
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MATT-CTR: Unleashing a Model-Agnostic Test-Time Paradigm for CTR Prediction with Confidence-Guided Inference Paths

DGX agent

arXiv:2510.08932v2 Announce Type: replace Abstract: Recently, a growing body of research has focused on either optimizing CTR model architectures to better model feature interactions or refining train

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

DGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

On the Generalization Capabilities, Design Choices and Limitations of Keypoint Imitation Learning

DGX agent

arXiv:2605.26649v1 Announce Type: new Abstract: RGB-based imitation learning requires many demonstrations to generalize to unseen objects or scenes, motivating research into intermediate representatio

tutorialsarxiv-cs-ro
27 May 2026
Model Releases

PRBench: A Standardized Probabilistic Robustness Benchmark

DGX agent

arXiv:2511.01724v3 Announce Type: replace Abstract: Deep learning models are notoriously vulnerable to imperceptible perturbations. Most existing research centers on adversarial robustness (AR), which

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Receipt Replay OOD: A Small Benchmark for Screen Replay Detection Under Domain Shift

DGX agent

arXiv:2605.26855v1 Announce Type: new Abstract: Public datasets such as DLC-2021, SynID, and KID34K have significantly contributed to research on presentation attack detection for identity documents,

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km

DGX agent

arXiv:2510.09606v2 Announce Type: replace Abstract: With the current surge in spatial reasoning explorations, researchers have made significant progress in understanding indoor scenes, but still strug

model-releasesarxiv-cs-cv
27 May 2026
Agents

AION: Next-Generation Tasks and Practical Harness for Time Series

DGX agent

arXiv:2605.25045v1 Announce Type: new Abstract: Time series research is moving beyond fixed forecasting benchmarks toward realistic tasks that combine prediction, contextual reasoning, tool use, and s

agentsarxiv-cs-ai
26 May 2026
Local Ai

Federated Learning over Human-Body Communication for On-Body Edge Intelligence: A Survey, Taxonomy, and BODYFED-HBC Scheduling Vignette

DGX agent

arXiv:2605.24062v1 Announce Type: cross Abstract: Human-body communication (HBC) is a promising physical substrate for wearable body-area networks because it can localize communication around the body

local-aiarxiv-cs-ai
26 May 2026
Model Releases

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

DGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

model-releasesarxiv-cs-lg
26 May 2026
Applications

Human-AI Collaboration in Science at Scale: A Global Large-scale Randomized Field Experiment

DGX agent

arXiv:2605.24180v1 Announce Type: cross Abstract: Collaboration is the defining mode of modern science, yet its core mechanism -- feedback -- remains hard to observe, difficult to scale, and unequally

applicationsarxiv-cs-ai
26 May 2026
Agents

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

DGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

agentsarxiv-cs-ai
26 May 2026
Safety

Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks

DGX agent

arXiv:2605.24217v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from research environments to production deployments, evaluating their performance against strict Service Lev

safetyarxiv-cs-ai
26 May 2026
Safety

Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI

DGX agent

arXiv:2605.23981v1 Announce Type: cross Abstract: Generative AI research increasingly confronts a shared problem: systems must sustain yet govern their own generative activity when uncertainty is high

safetyarxiv-cs-ai
26 May 2026
Model Releases

Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending

DGX agent

arXiv:2605.25574v1 Announce Type: cross Abstract: Concept erasure has emerged as a key research direction for ensuring safe and ethical image synthesis in Text-to-Image (T2I) models. While existing st

model-releasesarxiv-cs-ai
26 May 2026
Safety

QML-PipeGuard: Drift-Aware Behavioral Fingerprinting for Quantum Machine Learning Pipeline Integrity

DGX agent

arXiv:2605.25066v1 Announce Type: cross Abstract: Quantum machine learning (QML) is moving from research prototypes to deployed cloud services. As QML enters regulated industries, the integrity of the

safetyarxiv-cs-lg
26 May 2026
Safety

STaT: Resolving Shape Distortion in Non-Stationary Time Series via Tri-Modal Synergy

DGX agent

arXiv:2605.25943v1 Announce Type: new Abstract: Recent research in time series forecasting frequently investigates the integration of textual and visual modalities with numerical models to better navi

safetyarxiv-cs-lg
26 May 2026
Safety

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

DGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

safetyarxiv-cs-ro
26 May 2026
Model Releases

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

DGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Cultural Adaptation in Large Language Models for Political Discourse

DGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
Safety

Droneulator: A Portable UAV Simulator for Agricultural Workflows with RotorPy and Godot 4

DGX agent

arXiv:2605.23386v1 Announce Type: new Abstract: Agricultural UAV research requires simulators that integrate realistic 3D scenes, high-fidelity vehicle dynamics, and robotics middleware, while remaini

safetyarxiv-cs-ro
25 May 2026
Model Releases

How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness

DGX agent

arXiv:2605.23628v1 Announce Type: new Abstract: Multi-task benchmarks have become a central pillar of machine learning research, yet their growing influence has incentivised benchmark gaming -- strate

model-releasesarxiv-cs-lg
25 May 2026
Applications

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

DGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

applicationsarxiv-cs-cl
25 May 2026
Model Releases

Physiome-ODE: A Benchmark for Irregularly Sampled Multivariate Time Series Forecasting Based on Biological ODEs

DGX agent

arXiv:2502.07489v2 Announce Type: replace Abstract: State-of-the-art methods for forecasting irregularly sampled time series with missing values predominantly rely on just four datasets and a few smal

model-releasesarxiv-cs-lg
25 May 2026
Safety

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

DGX agent

arXiv:2605.23027v1 Announce Type: new Abstract: Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where r

safetyarxiv-cs-ro
25 May 2026
Safety

When Determinants Are Not Enough: Private Rare Switching

DGX agent

arXiv:2605.23131v1 Announce Type: new Abstract: In this note, I would like to share a small research moment where Codex helped me find the right way to adapt rare switching to the private setting. The

safetyarxiv-cs-lg
25 May 2026
Applications

Discovering Entity-Conditioned Lag Heterogeneity: A Lag-Gated Neural Audit Framework for Panel Time Series

DGX agent

arXiv:2605.21542v1 Announce Type: new Abstract: Country-level temporal panels are widely used in empirical analysis. Researchers often need to audit how different entities respond to historical signal

applicationsarxiv-cs-lg
23 May 2026
Model Releases

Does Slightly Mean Somewhat? Measuring Vague Intensity Words in LLM Numeric Actions

DGX agent

arXiv:2605.21827v1 Announce Type: new Abstract: Do language models preserve the ordinal meaning of intensity words when those words must produce numeric actions? I study a researcher-constructed scale

model-releasesarxiv-cs-cl
22 May 2026
Safety

AI-based Prediction of Independent Construction Safety Outcomes from Universal Attributes

DGX agent

arXiv:1908.05972v3 Announce Type: replace Abstract: This paper significantly improves on, and finishes to validate, an approach proposed in previous research in which safety outcomes were predicted fr

safetyarxiv-cs-lg
21 May 2026
Model Releases

Capability neq Interpretability: Human Interpretability of Vision Foundation Models

DGX agent

arXiv:2605.20337v1 Announce Type: new Abstract: How interpretable are the features of leading vision models? The question is increasingly pressing as these models move from research benchmarks into hi

model-releasesarxiv-cs-cv
21 May 2026
Safety

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

DGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

safetyarxiv-cs-cl
21 May 2026
Safety

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2605.21261v1 Announce Type: new Abstract: Training-free zero-shot composed image retrieval models are recently gaining increasing research interest due to their generalizability and flexibility

safetyarxiv-cs-cv
21 May 2026
Agents

A Logistic Regression Model to Predict Malaria Severity in Children

DGX agent

arXiv:2605.18900v1 Announce Type: cross Abstract: One of the main causes of death around the globe is malaria. Researchers have sought to develop predictive models for malaria outbreaks based on meteo

agentsarxiv-cs-lg
20 May 2026
Agents

Agent Security is a Systems Problem

DGX agent

arXiv:2605.18991v1 Announce Type: cross Abstract: We take the position that agent security must be approached as a systems problem: the AI model powering the agent must be treated as an untrusted comp

agentsarxiv-cs-ai
20 May 2026
Model Releases

AgentNLQ: A General-Purpose Agent for Natural Language to SQL

DGX agent

arXiv:2605.19010v1 Announce Type: new Abstract: Natural language to SQL (NL2SQL) conversion is an important problem for researchers and enterprises due to the ubiquitous importance of relational datab

model-releasesarxiv-cs-ai
20 May 2026
Agents

Cardiac fat segmentation using computed tomography and an image-to-image conditional generative adversarial neural network

DGX agent

arXiv:2605.20064v1 Announce Type: new Abstract: In recent years, research has highlighted the association between increased adipose tissue surrounding the human heart and elevated susceptibility to ca

agentsarxiv-cs-cv
20 May 2026
Agents

Causal Evidence that Language Models use Confidence to Drive Behavior

DGX agent

arXiv:2603.22161v2 Announce Type: replace Abstract: Metacognition -- assessing the quality of one's own cognitive performance -- guides adaptive behavior across species. Substantial research demonstra

agentsarxiv-cs-lg
20 May 2026
Applications

Data-Efficient Self-Supervised Algorithms for Fine-Grained Birdsong Analysis

DGX agent

arXiv:2511.12158v3 Announce Type: replace Abstract: Research in bioacoustics, neuroscience, and linguistics often uses birdsong as a proxy to acquire knowledge across diverse areas. This requires audi

applicationsarxiv-cs-lg
20 May 2026
Safety

Distributional AGI Safety

DGX agent

arXiv:2512.16856v2 Announce Type: replace Abstract: AI safety and alignment research has predominantly been focused on methods for safeguarding individual AI systems, resting on the assumption of an e

safetyarxiv-cs-ai
20 May 2026
Model Releases

GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation

DGX agent

arXiv:2605.19890v1 Announce Type: new Abstract: Test-time adaptation (TTA) enables a pre-trained model to adapt online to an unlabeled test stream under distribution shift. While most TTA research foc

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

JAXenstein: Accelerated Benchmarking for First-Person Environments

DGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Operationalizing Document AI: A Microservice Architecture for OCR and LLM Pipelines in Production

DGX agent

arXiv:2605.18818v1 Announce Type: new Abstract: Academic research tends to focus on new models for document understanding creating a wide gap in the literature between model definition and running mod

model-releasesarxiv-cs-ai
20 May 2026
Safety

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

DGX agent

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

safetyarxiv-cs-ai
20 May 2026
Model Releases

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

DGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

DGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

model-releasesarxiv-cs-ai
20 May 2026
Safety

Actionable World Representation

DGX agent

arXiv:2605.18743v1 Announce Type: new Abstract: Inspired by the emergent behaviors in large language models that generalized human intelligence, the research community is pursuing similar emergent cap

safetyarxiv-cs-ai
19 May 2026
Local Ai

CheckSupport: A Local LLM-Powered Tool for Automated Manuscript Submission Checklist Selection and Completion

DGX agent

arXiv:2605.16377v1 Announce Type: cross Abstract: Transparent and standardized reporting is essential for reproducible scientific research, yet adherence to reporting guidelines remains inconsistent b

local-aiarxiv-cs-ai
19 May 2026
← Previous
1…394395396397398…466
Next →