AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment

DGX agent

arXiv:2606.06667v1 Announce Type: new Abstract: The mechanisms behind LLMs' broad over-generalization beyond training examples remain unclear. Emergent misalignment (EM) offers a striking case study:

model-releasesarxiv-cs-cl
8 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning

DGX agent

arXiv:2606.06915v1 Announce Type: cross Abstract: Test-time compute (TTC) scaling has emerged as a powerful paradigm for improving large language model (LLM) reasoning by allocating additional compute

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Uncertainty-Aware LLM-Guided Policy Shaping for Sparse-Reward Reinforcement Learning

DGX agent

arXiv:2606.06673v1 Announce Type: new Abstract: Sparse rewards and heterogeneous task sequences remain persistent challenges in Reinforcement Learning (RL), often resulting in slow convergence, weak g

model-releasesarxiv-cs-lg
8 Jun 2026
Safety

What Do People Actually Want From AI? Mapping Preference Plurality

DGX agent

arXiv:2606.06674v1 Announce Type: new Abstract: Large Language Models (LLMs) are often fine-tuned through Reinforcement Learning from Human Feedback (RLHF) to align with people's preferences and value

safetyarxiv-cs-cl
8 Jun 2026
Model Releases

When to Think Deeply: Inhibitory Deliberation for LLM Reasoning

DGX agent

arXiv:2606.06745v1 Announce Type: new Abstract: Reasoning Large Language Models can improve problem-solving performance through deliberative inference, but invoking slow reasoning for every input is c

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs

DGX agent

arXiv:2601.12359v1 Announce Type: cross Abstract: Prompt injection attacks have become an increasing vulnerability for LLM applications, where adversarial prompts exploit indirect input channels such

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Agent-Orchestrated Adaptive RAG: A Comparative Study on Structured and Multi-Hop Retrieval

DGX agent

arXiv:2606.05658v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding their responses in external knowledge, but conventional pipeli

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Enhancing Software Engineering Through Closed-Loop Memory Optimization

DGX agent

arXiv:2606.05646v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled powerful software engineering (SE) agents capable of navigating complex codebases and resolving real-world i

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Evaluation of LLMs for Mathematical Formalization in Lean

DGX agent

arXiv:2606.05632v1 Announce Type: new Abstract: Within the past few years, the ability of Large Language Models (LLMs) to generate formal mathematical proofs has improved drastically. We provide a com

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

SagnacAssisted Enhanced OTDR for Distributed Acoustic Sensing: A Standardized Benchmark and Engineering Evaluation Framework

DGX agent

arXiv:2606.05754v1 Announce Type: cross Abstract: Phase-sensitive optical time-domain reflectometry (phi-OTDR) is widely used in large-scale distributed acoustic sensing (DAS) because it provides dist

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Synthetic Contrastive Reasoning for Multi-Table Q&A

DGX agent

arXiv:2606.05382v1 Announce Type: new Abstract: Multi-table question answering requires models to retrieve relevant evidence, link schemas, and perform compositional reasoning across relational tables

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems

DGX agent

arXiv:2606.05304v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models are typically organized around roles, pipelines, and turn schedules, while the content that age

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

When Tools Fail: Benchmarking Dynamic Replanning and Anomaly Recovery in LLM Agents

DGX agent

arXiv:2606.05806v1 Announce Type: new Abstract: Existing benchmarks evaluate Tool-Integrated Reasoning (TIR) in LLMs on idealized ''happy paths'', largely overlooking real-world tool failures. We intr

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Activation-Informed Pareto-Guided Low-Rank Compression for Efficient LLM/VLM

DGX agent

arXiv:2510.05544v2 Announce Type: replace Abstract: Large language models (LLM) and vision-language models (VLM) have achieved state-of-the-art performance, but they impose significant memory and comp

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Aligning Tree-Search Policies with Fixed Token Budgets in Test-Time Scaling of LLMs

DGX agent

arXiv:2602.09574v2 Announce Type: replace Abstract: Tree-search decoding is an effective form of test-time scaling for large language models (LLMs), but real-world deployment often imposes a fixed per

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Alignment Risks from Capability-Seeking RL Training

DGX agent

arXiv:2602.12124v2 Announce Type: replace-cross Abstract: While most AI alignment research focuses on preventing models from generating explicitly harmful content, a more subtle risk arises from capab

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

ArcANE: Do Role-Playing Language Agents Stay in Character at the Right Time?

DGX agent

arXiv:2606.05553v1 Announce Type: new Abstract: Role-playing language agents (RPLAs) should play characters whose values and behavior evolve as the story progresses, not maintain a fixed persona. Exis

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Asuka-Bench: Benchmarking Code Agents on Underspecified User Intent and Multi-Round Refinement

DGX agent

arXiv:2606.05920v1 Announce Type: cross Abstract: Existing code-generation benchmarks score a single mapping from a complete prompt to a one-shot output. However, real web development is different. Us

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

CoMoL: Efficient Mixture of LoRA Experts via Dynamic Core Space Merging

DGX agent

arXiv:2603.00573v2 Announce Type: replace Abstract: Large language models (LLMs) achieve remarkable performance on diverse downstream and domain-specific tasks via parameter-efficient fine-tuning (PEF

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Contextualized Prompting For Stance Detection On Social Media

DGX agent

arXiv:2606.06022v1 Announce Type: new Abstract: Stance detection on social media is challenging due to short, noisy, and context-dependent language. While large language models (LLMs) show zero-shot g

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Deep Learning-based 3D Oral Cavity Reconstruction Using 2D Intraoral Images

DGX agent

arXiv:2606.05998v1 Announce Type: new Abstract: Oral 3D modelling is one of the most essential stages in dentistry, and many different approaches, such as impression taking and intraoral scanning, are

researcharxiv-cs-cv
5 Jun 2026
Model Releases

DisasterBench: A Multimodal Benchmark for UAV-Based Disaster Response in Complex Environments

DGX agent

arXiv:2606.06217v1 Announce Type: new Abstract: When a disaster unfolds, responders must answer not only what is happening, but also why it is happening, what will happen next, and what to do now, oft

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

Domain-Conditioned Safety in Frontier Computer-Using Agents: A 793-Episode Browser Benchmark, a Coding-Domain Cross-Reference, and a Reproducibility Audit of Recent Red-Teaming

DGX agent

arXiv:2606.05233v1 Announce Type: cross Abstract: Recent computer-using-agent (CUA) red-teaming papers report prompt-injection attack success rates (ASR) of 42-98%, but these headline numbers cluster

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Facial-R1: Aligning Reasoning and Recognition for Facial Emotion Analysis

DGX agent

arXiv:2511.10254v2 Announce Type: replace Abstract: Facial Emotion Analysis (FEA) extends traditional facial emotion recognition by incorporating explainable, fine-grained reasoning. The task integrat

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD Maps

DGX agent

arXiv:2601.02730v3 Announce Type: replace Abstract: Visual localization on standard-definition (SD) maps has emerged as a promising low-cost and scalable solution for autonomous driving. However, exis

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

DGX agent

arXiv:2606.05624v1 Announce Type: new Abstract: Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop

model-releasesarxiv-cs-cv
5 Jun 2026
Safety

Latent Reasoning with Normalizing Flows

DGX agent

arXiv:2606.06447v1 Announce Type: new Abstract: Large language models often improve reasoning by generating explicit chain-of-thought (CoT), demonstrating the importance of intermediate computation. H

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

LiAuto-GeoX: Efficient Grounded Driving Transformer

DGX agent

arXiv:2606.05774v1 Announce Type: new Abstract: Dense 3D reconstruction has demonstrated immense potential for spatial understanding, yet its viability as a real-time, onboard representation for auton

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Multilingual Detection of Alzheimer's Disease from Speech: A Cross-Linguistic Transfer Learning Approach

DGX agent

arXiv:2606.05545v1 Announce Type: new Abstract: The development of multilingual Alzheimer's Disease Dementia (AD) detection models presents significant challenges due to the resource-intensive and tim

researcharxiv-cs-cl
5 Jun 2026
Model Releases

RAPTOR+: A Visually Grounded Vision-Language Framework to Improve Clinical Trust and Auditability in Automated Cancer Referral Processing

DGX agent

arXiv:2605.25956v2 Announce Type: replace Abstract: Urgent suspected colorectal cancer (CRC) referrals create operational bottlenecks because semi-structured clinical documents often require manual re

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

RedditPersona: A Modular Framework for Community-Conditioned LLM Adaptation from Reddit

DGX agent

arXiv:2606.06027v1 Announce Type: cross Abstract: Community-conditioned language model adaptation requires choices about data collection, community definition, and evaluation that are currently made i

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Stability vs. Manipulability: Evaluating Robustness Under Post-Decision Interaction in LLM Judges

DGX agent

arXiv:2606.05384v1 Announce Type: cross Abstract: LLM-as-judge evaluation is widely used in benchmarking pipelines, where model outputs are compared and ranked using automated evaluators. These pipeli

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Symb-xMIL: Symbolic Explanations for Multiple Instance Learning in Digital Pathology

DGX agent

arXiv:2606.06224v1 Announce Type: new Abstract: Explanations of multiple instance learning (MIL) models are widely used for validation and discovery in digital histopathology. Existing methods primari

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

TARPO: Token-Wise Latent-Explicit Reasoning via Action-Routing Policy Optimization

DGX agent

arXiv:2606.05859v1 Announce Type: new Abstract: Latent reasoning has emerged as a promising alternative to discrete Chain-of-Thought (CoT) in large language models (LLMs), enabling more expressive rea

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

V2V-Bench: A Comprehensive Benchmark for Video-to-Video Generation Evaluation

DGX agent

arXiv:2606.05665v1 Announce Type: new Abstract: Video-to-video (V2V) generation is difficult to evaluate because outputs must both follow editing instructions and preserve frame-level correspondence w

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

VZCrash: A Large-Scale IMU Dataset of Ego-Vehicle Crashes

DGX agent

arXiv:2606.06074v1 Announce Type: new Abstract: We introduce VZCrash, the largest publicly available dataset of real-world vehicle collision data featuring Inertial Measurement Unit (IMU) telemetry. T

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback

DGX agent

arXiv:2606.06113v1 Announce Type: new Abstract: Despite generating increasingly photorealistic images, text-to-image (T2I) models still exhibit localized, subtle, and structurally complex failures. Di

local-aiarxiv-cs-cv
5 Jun 2026
Local Ai

A Geometric View of Counterfactual Behavior: Interaction of Boundary Proximity and Local Support

DGX agent

arXiv:2606.04209v1 Announce Type: new Abstract: Counterfactual explanations seek small, semantically meaningful changes to an input that alter a model's prediction, and are widely used to interpret an

local-aiarxiv-cs-lg
4 Jun 2026
Model Releases

A Study of the Scale Invariant Signal to Distortion Ratio in Speech Separation with Noisy References

DGX agent

arXiv:2508.14623v2 Announce Type: replace-cross Abstract: This paper examines the implications of using the Scale-Invariant Signal-to-Distortion Ratio (SI-SDR) as both evaluation and training objectiv

model-releasesarxiv-cs-ai
4 Jun 2026
Hardware

AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2606.04484v1 Announce Type: new Abstract: We present AgentJet, a distributed swarm training framework for large language model (LLM) agent reinforcement learning. Unlike centralized frameworks t

hardwarearxiv-cs-ai
4 Jun 2026
Model Releases

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

DGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

COMBINER: Composed Image Retrieval Guided by Attribute-based Neighbor Relations

DGX agent

arXiv:2606.04604v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) represents a challenging retrieval task that targets locating specific images through multimodal inputs. Despite recent p

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Do Transformers Need Three Projections? Systematic Study of QKV Variants

DGX agent

arXiv:2606.04032v1 Announce Type: cross Abstract: Transformers have become the standard solution for various AI tasks, with the query, key, and value (QKV) attention formulation playing a central role

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

FLAGG: Flexible Autoregressive Graph Generation

DGX agent

arXiv:2606.05067v1 Announce Type: new Abstract: The Deep Graph Generation's panorama spans two extremes: one-shot and sequential models. The former generates nodes and edges jointly, while the latter

safetyarxiv-cs-lg
4 Jun 2026
Model Releases

GeM-NR: Geometry-Aware Multi-View Editing for Nonrigid Scene Changes

DGX agent

arXiv:2606.05142v1 Announce Type: cross Abstract: Recent developments in multi-view image editing with generative models have brought us a step closer toward general 3D content generation and customiz

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Impostor: An Agent-Curated Benchmark for Realistic AIGC Manipulation Localization

DGX agent

arXiv:2606.04545v1 Announce Type: new Abstract: Recent advances in generative image editing have improved the realism and controllability of localized image manipulation, raising new challenges for im

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space

DGX agent

arXiv:2604.04944v2 Announce Type: replace-cross Abstract: Multiple-choice questions (MCQs) are widely used to evaluate large language models (LLMs). However, LLMs remain vulnerable to the presence of

researcharxiv-cs-ai
4 Jun 2026
Model Releases

InstantRetouch: Efficient and High-Fidelity Instruction-Guided Image Retouching with Bilateral Space

DGX agent

arXiv:2606.05071v1 Announce Type: new Abstract: Language-guided photo retouching aims to adjust color and tone while preserving geometry and texture. Recently, diffusion-based retouching shows a super

model-releasesarxiv-cs-cv
4 Jun 2026
← Previous
1…416417418419420…1082
Next →