AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Safety

Alignment Makes Language Models Normative, Not Descriptive

DGX agent

arXiv:2603.17218v2 Announce Type: replace-cross Abstract: Post-training alignment optimizes language models to match human preference signals, but this objective is not equivalent to modeling observed

safetyarxiv-cs-ai
27 May 2026
Safety
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases

DGX agent

arXiv:2605.27355v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard method to align Large Language Models (LLMs) with human preferences. In this work, we

safetyarxiv-cs-ai
27 May 2026
Safety

Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines

DGX agent

arXiv:2605.26442v1 Announce Type: cross Abstract: Much of the alignment tuning literature is organized around optimization objectives, while the construction of alignment data is often treated implici

safetyarxiv-cs-ai
27 May 2026
Research

Amortized Factor Inference Networks for Posterior Inference

DGX agent

arXiv:2605.26419v1 Announce Type: new Abstract: Amortized inference promises fast test-time Bayesian inference, but existing methods are inherently tied to fixed models. Extending amortization to unse

researcharxiv-cs-lg
27 May 2026
Model Releases

An End-to-End Learning Approach for Solving Capacitated Location-Routing Problems

DGX agent

arXiv:2511.02525v2 Announce Type: replace-cross Abstract: The capacitated location-routing problems (CLRPs) are classical problems in combinatorial optimization, which require simultaneously making lo

model-releasesarxiv-cs-ai
27 May 2026
Research

An In-Vitro Study on Cross-Lingual Generalization in Language Models

DGX agent

arXiv:2605.26683v1 Announce Type: cross Abstract: Cross-lingual transfer in language models is difficult to study in natural corpora because lexical overlap, morphology, data imbalance, and tokenizati

researcharxiv-cs-ai
27 May 2026
Tutorials

An investigation of AI integration in sound designer workflows and experiences

DGX agent

arXiv:2605.27174v1 Announce Type: cross Abstract: Artificial intelligence is increasingly being integrated into professional audio production workflows, yet a gap persists between the tools developers

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification

DGX agent

arXiv:2503.21510v3 Announce Type: replace-cross Abstract: Ensuring that predictions of machine learning (ML) classification models are accompanied by uncertainty estimates is one of the main pillars o

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Anchor: Mitigating Artifact Drift in Agent Benchmark Generation

DGX agent

arXiv:2605.26321v1 Announce Type: new Abstract: AI agents are beginning to complete valuable, long-horizon business operations tasks, but training and evaluation environments for enterprise work still

model-releasesarxiv-cs-ai
27 May 2026
Research

AnchorDiff: Training-Free Concept Grounding for MM-DiTs via Anchor-Based Graph Propagation

DGX agent

arXiv:2605.26460v1 Announce Type: cross Abstract: Multi-Modal Diffusion Transformers (MM-DiTs) encode rich representations for training-free concept grounding, but existing attention-based methods oft

researcharxiv-cs-ai
27 May 2026
Research

Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model

DGX agent

arXiv:2602.07120v2 Announce Type: replace Abstract: Language models (LMs) tend to memorize portions of their training data and emit verbatim spans. When the underlying sources are sensitive or copyrig

researcharxiv-cs-cl
27 May 2026
Safety

Annotator Positionality as Signal: Psychometric Weighting for Anti-Autistic Ableism Detection

DGX agent

arXiv:2605.26397v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in decision-making tasks where they can amplify or suppress perspectives, raising concerns in high-

safetyarxiv-cs-ai
27 May 2026
Research

AnySurf: Any Surface Generation with Directed Edge

DGX agent

arXiv:2605.26149v1 Announce Type: cross Abstract: Open surface components prevail in real industrial 3D content and support rendering, physical simulation and geometric editing. Garments serve as a ty

researcharxiv-cs-cv
27 May 2026
Safety

Aperiodic and Low-Frequency Spectral Bias in Reconstruction based EEG Foundation Models

DGX agent

arXiv:2605.26434v1 Announce Type: cross Abstract: EEG foundation models, pre-trained on large-scale unlabelled EEG data, have emerged as a promising direction towards learning generalizable EEG repres

safetyarxiv-cs-ai
27 May 2026
Research

APEX: Amplitude Anchors and Phase Priors for Target-Scarce Higher-Frequency Wave Prediction

DGX agent

arXiv:2605.26732v1 Announce Type: new Abstract: Learning-based surrogates have become increasingly effective for wave-field prediction, and neural operators in particular have shown strong performance

researcharxiv-cs-lg
27 May 2026
Agents

APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2603.13853v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) connects large language models (LLMs) to external knowledge, but single-round retrieval is often insuffic

agentsarxiv-cs-ai
27 May 2026
Safety

Approximate Equivariance via Projection-based Regularisation

DGX agent

arXiv:2601.05028v2 Announce Type: replace Abstract: Equivariance is a powerful inductive bias in neural networks, improving generalisation and physical consistency. Recently, however, non-equivariant

safetyarxiv-cs-lg
27 May 2026
Model Releases

ARBITER: Reasoning Trajectory Basins and Majority Vote Failures in Test-Time Sampling

DGX agent

arXiv:2605.26172v1 Announce Type: new Abstract: When language models use test-time sampling, they generate multiple reasoning trajectories and select an answer by majority vote. We show that these tra

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation

DGX agent

arXiv:2605.26918v1 Announce Type: new Abstract: Video generation models (VGMs) are rapidly entering classrooms, yet existing benchmarks evaluate only perceptual quality, intrinsic faithfulness, generi

model-releasesarxiv-cs-cl
27 May 2026
Research

Assessing Per-Sample Membership Inference Vulnerability without Retraining

DGX agent

arXiv:2602.15919v2 Announce Type: replace-cross Abstract: Recent work in the privacy literature shows that sample-targeted membership inference attacks (MIAs) significantly outperform untargeted appro

researcharxiv-cs-ai
27 May 2026
Hardware

AssetGen: Deployable 3D Asset Generation at Interactive Speed

DGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

hardwarearxiv-cs-ai
27 May 2026
Safety

Athena: Enhancing Multimodal Reasoning with Data-efficient Process Reward Models

DGX agent

arXiv:2506.09532v5 Announce Type: replace-cross Abstract: We present Athena-PRM, a multimodal process reward model (PRM) designed to evaluate the reward score for each step in solving complex reasonin

safetyarxiv-cs-ai
27 May 2026
Agents

ATOM: Instantiating Budget-Controllable Multi-Agent Collaboration via Nucleus-Electron Hierarchy

DGX agent

arXiv:2605.26178v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems rely on optimized collaboration topologies to balance performance and communication costs. Howeve

agentsarxiv-cs-lg
27 May 2026
Research

Attenuation-Resilient Alternating Optimization for Laparoscopic Liver Landmark Detection

DGX agent

arXiv:2605.26630v1 Announce Type: new Abstract: Liver surface landmark detection is a fundamental prerequisite for anatomical guidance in laparoscopic liver surgery. However, it remains unreliable in

researcharxiv-cs-cv
27 May 2026
Model Releases

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

DGX agent

arXiv:2605.27025v1 Announce Type: new Abstract: Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematicall

model-releasesarxiv-cs-cl
27 May 2026
Safety

Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice

DGX agent

arXiv:2605.26559v1 Announce Type: cross Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks requi

safetyarxiv-cs-ai
27 May 2026
Applications

Augment Engineering: A Methodology for Multi-Tool AI Orchestration Across Professional Domains

DGX agent

arXiv:2605.26146v1 Announce Type: cross Abstract: Organizations increasingly deploy separate purpose-built AI tools across professional domains, often hiring domain specialists for each, recreating th

applicationsarxiv-cs-ai
27 May 2026
Model Releases

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

DGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

Automatic Layer Selection for Hallucination Detection

DGX agent

arXiv:2605.26366v1 Announce Type: new Abstract: Recent studies on hallucination detection have shown that hallucination-related signals are more strongly encoded in intermediate layers than in the fin

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

Axial-Centric Cross-Plane Attention for 3D Medical Image Classification

DGX agent

arXiv:2602.21636v2 Announce Type: replace Abstract: Abridged: Clinicians commonly interpret 3D medical images by examining multiple anatomical planes rather than relying on volumetric views. In clinic

model-releasesarxiv-cs-cv
27 May 2026
Safety

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

DGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

safetyarxiv-cs-cl
27 May 2026
Applications

Balancing Plasticity and Stability with Fast and Slow Successor Features

DGX agent

arXiv:2605.26357v1 Announce Type: new Abstract: A hallmark of intelligence is the ability to adapt in non-stationary environments, yet deep Reinforcement Learning (RL) agents often struggle in such se

applicationsarxiv-cs-lg
27 May 2026
Safety

BASIS: Batchwise Advantage Estimation from Single-Rollout Information Sharing for LLM Reasoning

DGX agent

arXiv:2605.27293v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become a standard recipe for improving the reasoning abilities of large language models. Existing alg

safetyarxiv-cs-lg
27 May 2026
Local Ai

BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting

DGX agent

arXiv:2605.27044v1 Announce Type: new Abstract: Early battery degradation trajectory forecasting (BDTF), which predicts the full-life state-of-health trajectory from early operational data, is critica

local-aiarxiv-cs-ai
27 May 2026
Model Releases

BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation

DGX agent

arXiv:2605.27067v1 Announce Type: new Abstract: Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegat

model-releasesarxiv-cs-cv
27 May 2026
Safety

Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility

DGX agent

arXiv:2603.03585v2 Announce Type: replace-cross Abstract: Misinformation is a growing societal threat, and susceptibility to misinformative claims varies across demographic groups due to differences i

safetyarxiv-cs-ai
27 May 2026
Model Releases

Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?

DGX agent

arXiv:2602.13626v3 Announce Type: replace Abstract: The expanding integration of Large Language Models (LLMs) into recommender systems poses critical challenges to evaluation reliability. This paper i

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

DGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback

DGX agent

arXiv:2509.21106v2 Announce Type: replace Abstract: Search-augmented large language models (LLMs) have advanced information-seeking tasks by integrating retrieval into generation, reducing users' cogn

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal

DGX agent

arXiv:2605.26772v1 Announce Type: new Abstract: Large reasoning models (LRMs) generate chain-of-thought (CoT) traces before producing final outputs, introducing a dynamic internal state that may compl

model-releasesarxiv-cs-ai
27 May 2026
Research

Beyond Binary: Speech Representations Across the Cognitive Score Hierarchy

DGX agent

arXiv:2605.27189v1 Announce Type: new Abstract: This study examines the relationship between speech representations and the hierarchical structure of cognitive assessment in mild cognitive impairment.

researcharxiv-cs-cl
27 May 2026
Safety

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation

DGX agent

arXiv:2601.03525v3 Announce Type: replace-cross Abstract: Effective reward design is a central challenge in Reinforcement Learning (RL) for code generation. Mainstream test-suite-level outcome rewards

safetyarxiv-cs-ai
27 May 2026
Research

Beyond Differences: Doubly Robust Meta-Learners for Ratio-Based Treatment Effects

DGX agent

arXiv:2605.26288v1 Announce Type: cross Abstract: When treatment effects are naturally expressed as ratios -- as in medicine, pricing, and marketing -- the ratio-based CATE au(x) = E[Y|W=1,X=x] / E[Y|

researcharxiv-cs-lg
27 May 2026
Safety

Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models

DGX agent

arXiv:2605.06213v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) today rests on fixed benchmarks that apply the same set of items to any model, producing ceiling and floor e

safetyarxiv-cs-ai
27 May 2026
Model Releases

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

DGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

model-releasesarxiv-cs-lg
27 May 2026
Safety

Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models

DGX agent

arXiv:2605.26491v1 Announce Type: cross Abstract: Preference optimization has emerged as an efficient alternative to online reinforcement learning from human feedback (RLHF) for aligning text-to-image

safetyarxiv-cs-cv
27 May 2026
Model Releases

Beyond Questions: Evaluating What Large Language Models (Actually) Know

DGX agent

arXiv:2605.26937v1 Announce Type: cross Abstract: Parametric knowledge in large language models (LLMs) is a cornerstone of their success, yet remains poorly understood. Existing knowledge benchmarks t

model-releasesarxiv-cs-ai
27 May 2026
Agents

Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent Systems

DGX agent

arXiv:2502.14321v3 Announce Type: replace-cross Abstract: Large language model-based multi-agent systems have recently gained significant attention due to their potential for complex, collaborative, a

agentsarxiv-cs-cl
27 May 2026
← Previous
1…730731732733734…1311
Next →