AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

Reinforcement Learning Enabled Adaptive Multi-Task Control for Bipedal Soccer Robots

DGX agent

arXiv:2604.19104v1 Announce Type: cross Abstract: Developing bipedal football robots in dynamiccombat environments presents challenges related to motionstability and deep coupling of multiple tasks, a

researcharxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports

DGX agent

arXiv:2604.19060v1 Announce Type: new Abstract: Accurate disease classification from radiology reports is essential for many applications. While supervised fine-tuning (SFT) of lightweight LLMs improv

safetyarxiv-cs-ai
22 Apr 2026
Tutorials

Relational AI in Education: Reciprocity, Participatory Design, and Indigenous Worldviews

DGX agent

arXiv:2604.19099v1 Announce Type: cross Abstract: Education is not merely the transmission of information or the optimisation of individual performance; it is a fundamentally social, constructive, and

tutorialsarxiv-cs-ai
22 Apr 2026
Agents

Remote Rowhammer Attack using Adversarial Observations on Federated Learning Clients

DGX agent

arXiv:2505.06335v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has the potential for simultaneous global learning amongst a large number of parallel agents, enabling emerging AI suc

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

RepIt: Steering Language Models with Concept-Specific Refusal Vectors

DGX agent

arXiv:2509.13281v5 Announce Type: replace Abstract: Current safety evaluations of language models rely on benchmark-based assessments that may miss localized vulnerabilities. We present RepIt, a simpl

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Resolving the Robustness-Precision Trade-off in Financial RAG through Hybrid Document-Routed Retrieval

DGX agent

arXiv:2603.26815v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems for financial document question answering typically follow a chunk-based paradigm: documents are

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

DGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

agentsarxiv-cs-ai
22 Apr 2026
Agents

Revac: A Social Deduction Reasoning Agent

DGX agent

arXiv:2604.19523v1 Announce Type: new Abstract: Social deduction games such as Mafia present a unique AI challenge: players must reason under uncertainty, interpret incomplete and intentionally mislea

agentsarxiv-cs-ai
22 Apr 2026
Safety

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

DGX agent

arXiv:2604.18757v1 Announce Type: cross Abstract: The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric f

safetyarxiv-cs-ai
22 Apr 2026
Research

Revisiting Catastrophic Forgetting in Continual Knowledge Graph Embedding

DGX agent

arXiv:2604.19401v1 Announce Type: cross Abstract: Knowledge Graph Embeddings (KGEs) support a wide range of downstream tasks over Knowledge Graphs (KGs). In practice, KGs evolve as new entities and fa

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Revisiting RaBitQ and TurboQuant: A Symmetric Comparison of Methods, Theory, and Experiments

DGX agent

arXiv:2604.19528v1 Announce Type: cross Abstract: This technical note revisits the relationship between RaBitQ and TurboQuant under a unified comparison framework. We compare the two methods in terms

model-releasesarxiv-cs-ai
22 Apr 2026
Research

RIFT: A RubrIc Failure Mode Taxonomy and Automated Diagnostics

DGX agent

arXiv:2604.01375v2 Announce Type: replace Abstract: Rubric-based evaluation is widely used in LLM benchmarks and training pipelines for open-ended, less verifiable tasks. While prior work has demonstr

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Right for the Wrong Reasons: Epistemic Regret Minimization for LLM Causal Reasoning

DGX agent

arXiv:2602.11675v3 Announce Type: replace Abstract: Large language models may answer causal questions correctly for the wrong reasons, substituting associational shortcuts P(Y|X) for the interventiona

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation

DGX agent

arXiv:2604.19092v1 Announce Type: cross Abstract: Recent advances in large-scale video world models have enabled increasingly realistic future prediction, raising the prospect of leveraging imagined v

model-releasesarxiv-cs-ai
22 Apr 2026
Applications

RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian

DGX agent

arXiv:2604.19593v1 Announce Type: cross Abstract: The importance of clear and correct text in legal documents cannot be understated, and, consequently, a grammatical error correction tool meant to ass

applicationsarxiv-cs-ai
22 Apr 2026
Applications

S2MAM: Semi-supervised Meta Additive Model for Robust Estimation and Variable Selection

DGX agent

arXiv:2604.19072v1 Announce Type: cross Abstract: Semi-supervised learning with manifold regularization is a classical framework for jointly learning from both labeled and unlabeled data, where the ke

applicationsarxiv-cs-ai
22 Apr 2026
Safety

Safety-Critical Contextual Control via Online Riemannian Optimization with World Models

DGX agent

arXiv:2604.19639v1 Announce Type: cross Abstract: Modern world models are becoming too complex to admit explicit dynamical descriptions. We study safety-critical contextual control, where a Planner mu

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models

DGX agent

arXiv:2604.19638v1 Announce Type: new Abstract: Multimodal Large Language Models are increasingly adopted as autonomous agents in interactive environments, yet their ability to proactively address saf

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

SAGE-32B: Agentic Reasoning via Iterative Distillation

DGX agent

arXiv:2601.04237v2 Announce Type: replace Abstract: We demonstrate SAGE-32B, a 32 billion parameter language model that focuses on agentic reasoning and long range planning tasks. Unlike chat models t

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

SAHM: A Benchmark for Arabic Financial and Shari'ah-Compliant Reasoning

DGX agent

arXiv:2604.19098v1 Announce Type: cross Abstract: English financial NLP has progressed rapidly through benchmarks for sentiment, document understanding, and financial question answering, while Arabic

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning

DGX agent

arXiv:2604.19048v1 Announce Type: cross Abstract: The combination of Mixture-of-Experts (MoE) and Low-Rank Adaptation (LoRA) has shown significant potential for enhancing the multi-task learning capab

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution

DGX agent

arXiv:2604.18982v1 Announce Type: new Abstract: Social intelligence, the ability to navigate complex interpersonal interactions, presents a fundamental challenge for language agents. Training such age

model-releasesarxiv-cs-ai
22 Apr 2026
Research

SCURank: Ranking Multiple Candidate Summaries with Summary Content Units for Enhanced Summarization

DGX agent

arXiv:2604.19185v1 Announce Type: cross Abstract: Small language models (SLMs), such as BART, can achieve summarization performance comparable to large language models (LLMs) via distillation. However

researcharxiv-cs-ai
22 Apr 2026
Safety

SEAT: Sparse Entity-Aware Tuning for Knowledge Adaptation while Preserving Epistemic Abstention

DGX agent

arXiv:2506.14387v3 Announce Type: replace Abstract: Adapting LLMs with new knowledge is increasingly important, but standard fine-tuning often erodes aligned epistemic abstention: the ability to ackno

safetyarxiv-cs-ai
22 Apr 2026
Research

See2Refine: Vision-Language Feedback Improves LLM-Based eHMI Action Designers

DGX agent

arXiv:2602.02063v2 Announce Type: replace-cross Abstract: Automated vehicles lack natural communication channels with other road users, making external Human-Machine Interfaces (eHMIs) essential for c

researcharxiv-cs-ai
22 Apr 2026
Local Ai

Self-Improving Tabular Language Models via Iterative Group Alignment

DGX agent

arXiv:2604.18966v1 Announce Type: cross Abstract: While language models have been adapted for tabular data generation, two fundamental limitations remain: (1) static fine-tuning produces models that c

local-aiarxiv-cs-ai
22 Apr 2026
Safety

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

DGX agent

arXiv:2604.18835v1 Announce Type: cross Abstract: We propose a scalable, multifactorial experimental framework that systematically probes LLM sensitivity to subtle semantic changes in pairwise documen

safetyarxiv-cs-ai
22 Apr 2026
Research

Sentipolis: Emotion-Aware Agents for Social Simulations

DGX agent

arXiv:2601.18027v2 Announce Type: replace Abstract: LLM agents are increasingly used for social simulation, yet emotion is often treated as a transient cue, causing emotional amnesia and weak long-hor

researcharxiv-cs-ai
22 Apr 2026
Model Releases

ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2604.19254v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) reduces the training cost of full-parameter fine-tuning for large language models (LLMs) by training only a sma

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Sherpa.ai Privacy-Preserving Multi-Party Entity Alignment without Intersection Disclosure for Noisy Identifiers

DGX agent

arXiv:2604.19219v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training among multiple parties without centralizing raw data. There are two main paradigms in FL:

safetyarxiv-cs-ai
22 Apr 2026
Research

SimDiff: Depth Pruning via Similarity and Difference

DGX agent

arXiv:2604.19520v1 Announce Type: new Abstract: Depth pruning improves the deployment efficiency of large language models (LLMs) by identifying and removing redundant layers. A widely accepted standar

researcharxiv-cs-ai
22 Apr 2026
Research

Skillful Global Ocean Emulation and the Role of Correlation-Aware Loss

DGX agent

arXiv:2604.18727v1 Announce Type: cross Abstract: Machine learning emulators have shown extraordinary skill in forecasting atmospheric states, and their application to global ocean dynamics offers sim

researcharxiv-cs-ai
22 Apr 2026
Model Releases

SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion

DGX agent

arXiv:2510.17925v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at code-related tasks but often struggle in realistic software repositories, where project-specific APIs an

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Speculative End-Turn Detector for Efficient Speech Chatbot Assistant

DGX agent

arXiv:2503.23439v2 Announce Type: replace-cross Abstract: Spoken dialogue systems powered by large language models have demonstrated remarkable abilities in understanding human speech and generating a

local-aiarxiv-cs-ai
22 Apr 2026
Hardware

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

DGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

hardwarearxiv-cs-ai
22 Apr 2026
Model Releases

SPRITE: From Static Mockups to Engine-Ready Game UI

DGX agent

arXiv:2604.18591v1 Announce Type: cross Abstract: Game UI implementation requires translating stylized mockups into interactive engine entities. However, current 'Screenshot-to-Code' tools often strug

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

ST-Prune: Training-Free Spatio-Temporal Token Pruning for Vision-Language Models in Autonomous Driving

DGX agent

arXiv:2604.19145v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become central to autonomous driving systems, yet their deployment is severely bottlenecked by the massive computat

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

StepFly: Agentic Troubleshooting Guide Automation for Incident Diagnosis

DGX agent

arXiv:2510.10074v2 Announce Type: replace Abstract: Effective incident management in large-scale IT systems relies on troubleshooting guides (TSGs), but their manual execution is slow and error-prone.

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Streamliners for Answer Set Programming

DGX agent

arXiv:2604.19251v1 Announce Type: cross Abstract: Streamliner constraints reduce the search space of combinatorial problems by ruling out portions of the solution space. We adapt the StreamLLM approac

researcharxiv-cs-ai
22 Apr 2026
Tutorials

TACENR: Task-Agnostic Contrastive Explanations for Node Representations

DGX agent

arXiv:2604.19372v1 Announce Type: cross Abstract: Graph representation learning has achieved notable success in encoding graph-structured data into latent vector spaces, enabling a wide range of downs

tutorialsarxiv-cs-ai
22 Apr 2026
Research

Tadabur: A Large-Scale Quran Audio Dataset

DGX agent

arXiv:2604.18932v1 Announce Type: cross Abstract: Despite growing interest in Quranic data research, existing Quran datasets remain limited in both scale and diversity. To address this gap, we present

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

DGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

DGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning

DGX agent

arXiv:2601.18296v2 Announce Type: replace-cross Abstract: Temporal Knowledge Graph Question Answering (TKGQA) is inherently challenging, as it requires sophisticated reasoning over dynamic facts with

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Temporal UI State Inconsistency in Desktop GUI Agents: Formalizing and Defending Against TOCTOU Attacks on Computer-Use Agents

DGX agent

arXiv:2604.18860v1 Announce Type: cross Abstract: GUI agents that control desktop computers via screenshot-and-click loops introduce a new class of vulnerability: the observation-to-action gap (mean 6

researcharxiv-cs-ai
22 Apr 2026
Hardware

Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters

DGX agent

arXiv:2509.18831v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have significantly improved image and video synthesis. In addition, several concept control methods have b

hardwarearxiv-cs-ai
22 Apr 2026
Agents

TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction

DGX agent

arXiv:2602.06400v2 Announce Type: replace-cross Abstract: The prediction of 3D semantic occupancy enables autonomous vehicles (AVs) to perceive the fine-grained geometric and semantic scene structure

agentsarxiv-cs-ai
22 Apr 2026
Research

The Cost of Relaxation: Evaluating the Error in Convex Neural Network Verification

DGX agent

arXiv:2604.18728v1 Announce Type: cross Abstract: Many neural network (NN) verification systems represent the network's input-output relation as a constraint program. Sound and complete, representatio

researcharxiv-cs-ai
22 Apr 2026
← Previous
1…399400401402403…443
Next →