AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Safety

Alignment Makes Language Models Normative, Not Descriptive

DGX agent

arXiv:2603.17218v2 Announce Type: replace-cross Abstract: Post-training alignment optimizes language models to match human preference signals, but this objective is not equivalent to modeling observed

safetyarxiv-cs-ai
27 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases

DGX agent

arXiv:2605.27355v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard method to align Large Language Models (LLMs) with human preferences. In this work, we

safetyarxiv-cs-ai
27 May 2026
Safety

Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines

DGX agent

arXiv:2605.26442v1 Announce Type: cross Abstract: Much of the alignment tuning literature is organized around optimization objectives, while the construction of alignment data is often treated implici

safetyarxiv-cs-ai
27 May 2026
Research

Amortized Factor Inference Networks for Posterior Inference

DGX agent

arXiv:2605.26419v1 Announce Type: new Abstract: Amortized inference promises fast test-time Bayesian inference, but existing methods are inherently tied to fixed models. Extending amortization to unse

researcharxiv-cs-lg
27 May 2026
Model Releases

An End-to-End Learning Approach for Solving Capacitated Location-Routing Problems

DGX agent

arXiv:2511.02525v2 Announce Type: replace-cross Abstract: The capacitated location-routing problems (CLRPs) are classical problems in combinatorial optimization, which require simultaneously making lo

model-releasesarxiv-cs-ai
27 May 2026
Research

An In-Vitro Study on Cross-Lingual Generalization in Language Models

DGX agent

arXiv:2605.26683v1 Announce Type: cross Abstract: Cross-lingual transfer in language models is difficult to study in natural corpora because lexical overlap, morphology, data imbalance, and tokenizati

researcharxiv-cs-ai
27 May 2026
Tutorials

An investigation of AI integration in sound designer workflows and experiences

DGX agent

arXiv:2605.27174v1 Announce Type: cross Abstract: Artificial intelligence is increasingly being integrated into professional audio production workflows, yet a gap persists between the tools developers

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification

DGX agent

arXiv:2503.21510v3 Announce Type: replace-cross Abstract: Ensuring that predictions of machine learning (ML) classification models are accompanied by uncertainty estimates is one of the main pillars o

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Anchor: Mitigating Artifact Drift in Agent Benchmark Generation

DGX agent

arXiv:2605.26321v1 Announce Type: new Abstract: AI agents are beginning to complete valuable, long-horizon business operations tasks, but training and evaluation environments for enterprise work still

model-releasesarxiv-cs-ai
27 May 2026
Research

AnchorDiff: Training-Free Concept Grounding for MM-DiTs via Anchor-Based Graph Propagation

DGX agent

arXiv:2605.26460v1 Announce Type: cross Abstract: Multi-Modal Diffusion Transformers (MM-DiTs) encode rich representations for training-free concept grounding, but existing attention-based methods oft

researcharxiv-cs-ai
27 May 2026
Research

Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model

DGX agent

arXiv:2602.07120v2 Announce Type: replace Abstract: Language models (LMs) tend to memorize portions of their training data and emit verbatim spans. When the underlying sources are sensitive or copyrig

researcharxiv-cs-cl
27 May 2026
Safety

Annotator Positionality as Signal: Psychometric Weighting for Anti-Autistic Ableism Detection

DGX agent

arXiv:2605.26397v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in decision-making tasks where they can amplify or suppress perspectives, raising concerns in high-

safetyarxiv-cs-ai
27 May 2026
Research

AnySurf: Any Surface Generation with Directed Edge

DGX agent

arXiv:2605.26149v1 Announce Type: cross Abstract: Open surface components prevail in real industrial 3D content and support rendering, physical simulation and geometric editing. Garments serve as a ty

researcharxiv-cs-cv
27 May 2026
Safety

Aperiodic and Low-Frequency Spectral Bias in Reconstruction based EEG Foundation Models

DGX agent

arXiv:2605.26434v1 Announce Type: cross Abstract: EEG foundation models, pre-trained on large-scale unlabelled EEG data, have emerged as a promising direction towards learning generalizable EEG repres

safetyarxiv-cs-ai
27 May 2026
Research

APEX: Amplitude Anchors and Phase Priors for Target-Scarce Higher-Frequency Wave Prediction

DGX agent

arXiv:2605.26732v1 Announce Type: new Abstract: Learning-based surrogates have become increasingly effective for wave-field prediction, and neural operators in particular have shown strong performance

researcharxiv-cs-lg
27 May 2026
Agents

APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2603.13853v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) connects large language models (LLMs) to external knowledge, but single-round retrieval is often insuffic

agentsarxiv-cs-ai
27 May 2026
Safety

Approximate Equivariance via Projection-based Regularisation

DGX agent

arXiv:2601.05028v2 Announce Type: replace Abstract: Equivariance is a powerful inductive bias in neural networks, improving generalisation and physical consistency. Recently, however, non-equivariant

safetyarxiv-cs-lg
27 May 2026
Model Releases

ARBITER: Reasoning Trajectory Basins and Majority Vote Failures in Test-Time Sampling

DGX agent

arXiv:2605.26172v1 Announce Type: new Abstract: When language models use test-time sampling, they generate multiple reasoning trajectories and select an answer by majority vote. We show that these tra

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation

DGX agent

arXiv:2605.26918v1 Announce Type: new Abstract: Video generation models (VGMs) are rapidly entering classrooms, yet existing benchmarks evaluate only perceptual quality, intrinsic faithfulness, generi

model-releasesarxiv-cs-cl
27 May 2026
Research

Assessing Per-Sample Membership Inference Vulnerability without Retraining

DGX agent

arXiv:2602.15919v2 Announce Type: replace-cross Abstract: Recent work in the privacy literature shows that sample-targeted membership inference attacks (MIAs) significantly outperform untargeted appro

researcharxiv-cs-ai
27 May 2026
Hardware

AssetGen: Deployable 3D Asset Generation at Interactive Speed

DGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

hardwarearxiv-cs-ai
27 May 2026
Safety

Athena: Enhancing Multimodal Reasoning with Data-efficient Process Reward Models

DGX agent

arXiv:2506.09532v5 Announce Type: replace-cross Abstract: We present Athena-PRM, a multimodal process reward model (PRM) designed to evaluate the reward score for each step in solving complex reasonin

safetyarxiv-cs-ai
27 May 2026
Agents

ATOM: Instantiating Budget-Controllable Multi-Agent Collaboration via Nucleus-Electron Hierarchy

DGX agent

arXiv:2605.26178v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems rely on optimized collaboration topologies to balance performance and communication costs. Howeve

agentsarxiv-cs-lg
27 May 2026
Research

Attenuation-Resilient Alternating Optimization for Laparoscopic Liver Landmark Detection

DGX agent

arXiv:2605.26630v1 Announce Type: new Abstract: Liver surface landmark detection is a fundamental prerequisite for anatomical guidance in laparoscopic liver surgery. However, it remains unreliable in

researcharxiv-cs-cv
27 May 2026
Model Releases

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

DGX agent

arXiv:2605.27025v1 Announce Type: new Abstract: Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematicall

model-releasesarxiv-cs-cl
27 May 2026
Safety

Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice

DGX agent

arXiv:2605.26559v1 Announce Type: cross Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks requi

safetyarxiv-cs-ai
27 May 2026
Applications

Augment Engineering: A Methodology for Multi-Tool AI Orchestration Across Professional Domains

DGX agent

arXiv:2605.26146v1 Announce Type: cross Abstract: Organizations increasingly deploy separate purpose-built AI tools across professional domains, often hiring domain specialists for each, recreating th

applicationsarxiv-cs-ai
27 May 2026
Model Releases

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

DGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

Automatic Layer Selection for Hallucination Detection

DGX agent

arXiv:2605.26366v1 Announce Type: new Abstract: Recent studies on hallucination detection have shown that hallucination-related signals are more strongly encoded in intermediate layers than in the fin

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

Axial-Centric Cross-Plane Attention for 3D Medical Image Classification

DGX agent

arXiv:2602.21636v2 Announce Type: replace Abstract: Abridged: Clinicians commonly interpret 3D medical images by examining multiple anatomical planes rather than relying on volumetric views. In clinic

model-releasesarxiv-cs-cv
27 May 2026
Safety

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

DGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

safetyarxiv-cs-cl
27 May 2026
Applications

Balancing Plasticity and Stability with Fast and Slow Successor Features

DGX agent

arXiv:2605.26357v1 Announce Type: new Abstract: A hallmark of intelligence is the ability to adapt in non-stationary environments, yet deep Reinforcement Learning (RL) agents often struggle in such se

applicationsarxiv-cs-lg
27 May 2026
Safety

BASIS: Batchwise Advantage Estimation from Single-Rollout Information Sharing for LLM Reasoning

DGX agent

arXiv:2605.27293v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become a standard recipe for improving the reasoning abilities of large language models. Existing alg

safetyarxiv-cs-lg
27 May 2026
Local Ai

BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting

DGX agent

arXiv:2605.27044v1 Announce Type: new Abstract: Early battery degradation trajectory forecasting (BDTF), which predicts the full-life state-of-health trajectory from early operational data, is critica

local-aiarxiv-cs-ai
27 May 2026
Model Releases

BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation

DGX agent

arXiv:2605.27067v1 Announce Type: new Abstract: Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegat

model-releasesarxiv-cs-cv
27 May 2026
Safety

Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility

DGX agent

arXiv:2603.03585v2 Announce Type: replace-cross Abstract: Misinformation is a growing societal threat, and susceptibility to misinformative claims varies across demographic groups due to differences i

safetyarxiv-cs-ai
27 May 2026
Model Releases

Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?

DGX agent

arXiv:2602.13626v3 Announce Type: replace Abstract: The expanding integration of Large Language Models (LLMs) into recommender systems poses critical challenges to evaluation reliability. This paper i

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

DGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback

DGX agent

arXiv:2509.21106v2 Announce Type: replace Abstract: Search-augmented large language models (LLMs) have advanced information-seeking tasks by integrating retrieval into generation, reducing users' cogn

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal

DGX agent

arXiv:2605.26772v1 Announce Type: new Abstract: Large reasoning models (LRMs) generate chain-of-thought (CoT) traces before producing final outputs, introducing a dynamic internal state that may compl

model-releasesarxiv-cs-ai
27 May 2026
Research

Beyond Binary: Speech Representations Across the Cognitive Score Hierarchy

DGX agent

arXiv:2605.27189v1 Announce Type: new Abstract: This study examines the relationship between speech representations and the hierarchical structure of cognitive assessment in mild cognitive impairment.

researcharxiv-cs-cl
27 May 2026
Safety

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation

DGX agent

arXiv:2601.03525v3 Announce Type: replace-cross Abstract: Effective reward design is a central challenge in Reinforcement Learning (RL) for code generation. Mainstream test-suite-level outcome rewards

safetyarxiv-cs-ai
27 May 2026
Research

Beyond Differences: Doubly Robust Meta-Learners for Ratio-Based Treatment Effects

DGX agent

arXiv:2605.26288v1 Announce Type: cross Abstract: When treatment effects are naturally expressed as ratios -- as in medicine, pricing, and marketing -- the ratio-based CATE au(x) = E[Y|W=1,X=x] / E[Y|

researcharxiv-cs-lg
27 May 2026
Safety

Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models

DGX agent

arXiv:2605.06213v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) today rests on fixed benchmarks that apply the same set of items to any model, producing ceiling and floor e

safetyarxiv-cs-ai
27 May 2026
Model Releases

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

DGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

model-releasesarxiv-cs-lg
27 May 2026
Safety

Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models

DGX agent

arXiv:2605.26491v1 Announce Type: cross Abstract: Preference optimization has emerged as an efficient alternative to online reinforcement learning from human feedback (RLHF) for aligning text-to-image

safetyarxiv-cs-cv
27 May 2026
Model Releases

Beyond Questions: Evaluating What Large Language Models (Actually) Know

DGX agent

arXiv:2605.26937v1 Announce Type: cross Abstract: Parametric knowledge in large language models (LLMs) is a cornerstone of their success, yet remains poorly understood. Existing knowledge benchmarks t

model-releasesarxiv-cs-ai
27 May 2026
Agents

Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent Systems

DGX agent

arXiv:2502.14321v3 Announce Type: replace-cross Abstract: Large language model-based multi-agent systems have recently gained significant attention due to their potential for complex, collaborative, a

agentsarxiv-cs-cl
27 May 2026
← Previous
1…730731732733734…1311
Next →