AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Research

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs

DGX agent

arXiv:2602.11528v2 Announce Type: replace-cross Abstract: Recent studies have shown that large language models (LLMs) can infer private user attributes (e.g., age, location, gender) from user-generate

researcharxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts

DGX agent

arXiv:2604.18473v1 Announce Type: new Abstract: Extending a fully post-trained language model with new domain capabilities is fundamentally limited by monolithic training paradigms: retraining from sc

safetyarxiv-cs-lg
21 Apr 2026
Local Ai

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts

DGX agent

arXiv:2604.16542v1 Announce Type: cross Abstract: Safety guardrails have become an active area of research in AI safety, aimed at ensuring the appropriate behavior of large language models (LLMs). How

local-aiarxiv-cs-cl
21 Apr 2026
Local Ai

VocabTailor: Dynamic Vocabulary Selection for Downstream Tasks in Small Language Models

DGX agent

arXiv:2508.15229v3 Announce Type: replace Abstract: Small Language Models (SLMs) provide computational advantages in resource-constrained environments, yet memory limitations remain a critical bottlen

local-aiarxiv-cs-cl
21 Apr 2026
Local Ai

Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

DGX agent

arXiv:2604.17422v1 Announce Type: new Abstract: Long video understanding remains a formidable challenge for Multimodal Large Language Models (MLLMs) due to the prohibitive computational cost of proces

local-aiarxiv-cs-cv
21 Apr 2026
Local Ai

Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models

DGX agent

arXiv:2604.15609v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) for black-box models accessible only via APIs remains a largely unexplored challenge. Existing approaches such as post-hoc

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

An Information-Geometric Approach to Artificial Curiosity

DGX agent

arXiv:2504.06355v2 Announce Type: replace Abstract: Learning in environments with sparse rewards remains a fundamental challenge in reinforcement learning. Artificial curiosity addresses this limitati

model-releasesarxiv-cs-lg
20 Apr 2026
Safety

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

DGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

safetyarxiv-cs-cv
20 Apr 2026
Research

CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics

DGX agent

arXiv:2604.15647v1 Announce Type: new Abstract: Measuring the quality of public deliberation requires evaluating not only civility or argument structure, but also the informational progress of a conve

researcharxiv-cs-cl
20 Apr 2026
Model Releases

COEVO: Co-Evolutionary Framework for Joint Functional Correctness and PPA Optimization in LLM-Based RTL Generation

DGX agent

arXiv:2604.15001v2 Announce Type: replace Abstract: LLM-based RTL code generation methods increasingly target both functional correctness and PPA quality, yet existing approaches universally decouple

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Information-Consistent Language Model Recommendations through Group Relative Policy Optimization

DGX agent

arXiv:2512.12858v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in business-critical domains such as finance, education, healthcare, and customer suppo

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

LaMSUM: Amplifying Voices Against Harassment through LLM Guided Extractive Summarization of User Incident Reports

DGX agent

arXiv:2406.15809v5 Announce Type: replace Abstract: Citizen reporting platforms help the public and authorities stay informed about sexual harassment incidents. However, the high volume of data shared

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment

DGX agent

arXiv:2604.15757v1 Announce Type: new Abstract: This research note identifies a previously overlooked distinction between multi-objective reinforcement learning (MORL), and more conventional single-ob

safetyarxiv-cs-lg
20 Apr 2026
Hardware

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

DGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

hardwarearxiv-cs-ro
20 Apr 2026
Local Ai

Prompt-Driven Code Summarization: A Systematic Literature Review

DGX agent

arXiv:2604.15385v1 Announce Type: cross Abstract: Software documentation is essential for program comprehension, developer onboarding, code review, and long-term maintenance. Yet producing quality doc

local-aiarxiv-cs-lg
20 Apr 2026
Safety

Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability

DGX agent

arXiv:2604.16106v1 Announce Type: cross Abstract: The Project of AI is a world-building endeavor, wherein those who fund and develop AI systems both operate through and seek to sustain networks of pow

safetyarxiv-cs-ai
20 Apr 2026
Local Ai

Revisiting the Uniform Information Density Hypothesis in LLM Reasoning

DGX agent

arXiv:2510.06953v3 Announce Type: replace Abstract: The Uniform Information Density (UID) hypothesis proposes that effective communication is achieved by maintaining a stable flow of information. In t

local-aiarxiv-cs-ai
20 Apr 2026
Safety

Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration

DGX agent

arXiv:2503.05578v4 Announce Type: replace Abstract: Pose estimation-guided unseen object 6-DoF robotic manipulation is a key task in robotics. However, the scalability of current pose estimation metho

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

Seed1.8 Model Card: Towards Generalized Real-World Agency

DGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Skill-RAG: Failure-State-Aware Retrieval Augmentation via Hidden-State Probing and Skill Routing

DGX agent

arXiv:2604.15771v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a foundational paradigm for grounding large language models in external knowledge. While adaptive re

safetyarxiv-cs-cl
20 Apr 2026
Hardware

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

DGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

hardwarearxiv-cs-ai
20 Apr 2026
Local Ai

Two-Stage Framework for Efficient UAV-Based Wildfire Video Analysis with Adaptive Compression and Fire Source Detection

DGX agent

arXiv:2508.16739v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) have become increasingly important in disaster emergency response by facilitating aerial video analysis. Due to the

local-aiarxiv-cs-cv
20 Apr 2026
Safety

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

DGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

Why Fine-Tuning Encourages Hallucinations and How to Fix It

DGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot

DGX agent

arXiv:2604.13940v1 Announce Type: new Abstract: Scientific peer review faces mounting strain as submission volumes surge, making it increasingly difficult to sustain review quality, consistency, and t

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Benchmarking Classical Coverage Path Planning Heuristics on Irregular Hexagonal Grids for Maritime Coverage Scenarios

DGX agent

arXiv:2604.15202v1 Announce Type: new Abstract: Coverage path planning on irregular hexagonal grids is relevant to maritime surveillance, search and rescue and environmental monitoring, yet classical

model-releasesarxiv-cs-ro
17 Apr 2026
Safety

Beyond Importance Sampling: Rejection-Gated Policy Optimization

DGX agent

arXiv:2604.14895v1 Announce Type: new Abstract: We propose a new perspective on policy optimization: rather than reweighting all samples by their importance ratios, an optimizer should select which sa

safetyarxiv-cs-lg
17 Apr 2026
Safety

Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counseling

DGX agent

arXiv:2604.15124v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) is central to diabetes care, but explaining CGM patterns clearly and empathetically remains time-intensive. Evidence

safetyarxiv-cs-cl
17 Apr 2026
Safety

BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation

DGX agent

arXiv:2510.20151v2 Announce Type: replace Abstract: Structured texts refer to texts containing structured elements beyond plain texts, such as code snippets and placeholders. Such structured texts inc

safetyarxiv-cs-cl
17 Apr 2026
Safety

Controllable Video Object Insertion via Multiview Priors

DGX agent

arXiv:2604.14556v1 Announce Type: new Abstract: Video object insertion is a critical task for dynamically inserting new objects into existing environments. Previous video generation methods focus prim

safetyarxiv-cs-cv
17 Apr 2026
Hardware

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

DGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

hardwarearxiv-cs-cl
17 Apr 2026
Safety

From Plausible to Causal: Counterfactual Semantics for Policy Evaluation in Simulated Online Communities

DGX agent

arXiv:2604.03920v2 Announce Type: replace Abstract: LLM-based social simulations can generate believable community interactions, enabling ``policy wind tunnels'' where governance interventions are tes

safetyarxiv-cs-cl
17 Apr 2026
Safety

GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

DGX agent

arXiv:2604.14258v1 Announce Type: cross Abstract: Large language models are typically post-trained using supervised fine-tuning (SFT) and reinforcement learning (RL), yet effectively unifying efficien

safetyarxiv-cs-lg
17 Apr 2026
Hardware

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding

DGX agent

arXiv:2601.14724v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated significant improvement in offline video understanding. Howe

hardwarearxiv-cs-cl
17 Apr 2026
Safety

IROSA: Interactive Robot Skill Adaptation using Natural Language

DGX agent

arXiv:2603.03897v3 Announce Type: replace-cross Abstract: Foundation models have demonstrated impressive capabilities across diverse domains, while imitation learning provides principled methods for r

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems

DGX agent

arXiv:2604.14799v1 Announce Type: new Abstract: Effective abstention (EA), recognizing evidence insufficiency and refraining from answering, is critical for reliable multimodal systems. Yet existing e

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality

DGX agent

arXiv:2506.19807v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly slow-thinking models, often exhibit severe hallucination, outputting incorrect content due to an in

tutorialsarxiv-cs-cl
17 Apr 2026
Hardware

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers

DGX agent

arXiv:2509.23638v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models face memory and PCIe latency bottlenecks when deployed on commodity hardware. Offloading expert weights to CPU memor

hardwarearxiv-cs-lg
17 Apr 2026
Safety

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

DGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

safetyarxiv-cs-lg
17 Apr 2026
Model Releases

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning

DGX agent

arXiv:2604.14922v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a critical driver for enhancing the reasoning capabilities of Large Language Models (LLMs). While recent ad

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Meituan Merchant Business Diagnosis via Policy-Guided Dual-Process User Simulation

DGX agent

arXiv:2604.15190v1 Announce Type: cross Abstract: Simulating group-level user behavior enables scalable counterfactual evaluation of merchant strategies without costly online experiments. However, bui

safetyarxiv-cs-cl
17 Apr 2026
Hardware

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

DGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

hardwarearxiv-cs-cl
17 Apr 2026
Safety

Neuro-Symbolic AI for Cybersecurity: State of the Art, Challenges, and Opportunities

DGX agent

arXiv:2509.06921v2 Announce Type: replace-cross Abstract: Cybersecurity demands both rapid pattern recognition and deliberative reasoning, yet purely neural or purely symbolic approaches each address

safetyarxiv-cs-ai
17 Apr 2026
Model Releases

Prompt Optimization Is a Coin Flip: Diagnosing When It Helps in Compound AI Systems

DGX agent

arXiv:2604.14585v1 Announce Type: cross Abstract: Prompt optimization in compound AI systems is statistically indistinguishable from a coin flip: across 72 optimization runs on Claude Haiku (6 methods

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Rethinking Patient Education as Multi-turn Multi-modal Interaction

DGX agent

arXiv:2604.14656v1 Announce Type: cross Abstract: Most medical multimodal benchmarks focus on static tasks such as image question answering, report generation, and plain-language rewriting. Patient ed

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

RL-STPA: Adapting System-Theoretic Hazard Analysis for Safety-Critical Reinforcement Learning

DGX agent

arXiv:2604.15201v1 Announce Type: new Abstract: As reinforcement learning (RL) deployments expand into safety-critical domains, existing evaluation methods fail to systematically identify hazards aris

safetyarxiv-cs-lg
17 Apr 2026
Safety

Scouting By Reward: VLM-TO-IRL-Driven Player Selection For Esports

DGX agent

arXiv:2604.14474v1 Announce Type: new Abstract: Traditional esports scouting workflows rely heavily on manual video review and aggregate performance metrics, which often fail to capture the nuanced de

safetyarxiv-cs-lg
17 Apr 2026
Safety

Separation is Optimal for LQR under Intermittent Feedback

DGX agent

arXiv:2603.27833v2 Announce Type: cross Abstract: In this work, we first prove that the separation principle holds for communication-constrained LQR problems under i.i.d. zero-mean disturbances with a

safetyarxiv-cs-ro
17 Apr 2026
← Previous
1…197198199200201…233
Next →