AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
22 Apr 2026

Reinforcement Learning Enabled Adaptive Multi-Task Control for Bipedal Soccer Robots

ResearchDGX agent

arXiv:2604.19104v1 Announce Type: cross Abstract: Developing bipedal football robots in dynamiccombat environments presents challenges related to motionstability and deep coupling of multiple tasks, a

Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports

SafetyDGX agent

arXiv:2604.19060v1 Announce Type: new Abstract: Accurate disease classification from radiology reports is essential for many applications. While supervised fine-tuning (SFT) of lightweight LLMs improv

Relational AI in Education: Reciprocity, Participatory Design, and Indigenous Worldviews

TutorialsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.19099v1 Announce Type: cross Abstract: Education is not merely the transmission of information or the optimisation of individual performance; it is a fundamentally social, constructive, and

Remote Rowhammer Attack using Adversarial Observations on Federated Learning Clients

AgentsDGX agent

arXiv:2505.06335v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has the potential for simultaneous global learning amongst a large number of parallel agents, enabling emerging AI suc

RepIt: Steering Language Models with Concept-Specific Refusal Vectors

Model ReleasesDGX agent

arXiv:2509.13281v5 Announce Type: replace Abstract: Current safety evaluations of language models rely on benchmark-based assessments that may miss localized vulnerabilities. We present RepIt, a simpl

Resolving the Robustness-Precision Trade-off in Financial RAG through Hybrid Document-Routed Retrieval

Model ReleasesDGX agent

arXiv:2603.26815v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems for financial document question answering typically follow a chunk-based paradigm: documents are

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

AgentsDGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

Revac: A Social Deduction Reasoning Agent

AgentsDGX agent

arXiv:2604.19523v1 Announce Type: new Abstract: Social deduction games such as Mafia present a unique AI challenge: players must reason under uncertainty, interpret incomplete and intentionally mislea

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

SafetyDGX agent

arXiv:2604.18757v1 Announce Type: cross Abstract: The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric f

Revisiting Catastrophic Forgetting in Continual Knowledge Graph Embedding

ResearchDGX agent

arXiv:2604.19401v1 Announce Type: cross Abstract: Knowledge Graph Embeddings (KGEs) support a wide range of downstream tasks over Knowledge Graphs (KGs). In practice, KGs evolve as new entities and fa

Revisiting RaBitQ and TurboQuant: A Symmetric Comparison of Methods, Theory, and Experiments

Model ReleasesDGX agent

arXiv:2604.19528v1 Announce Type: cross Abstract: This technical note revisits the relationship between RaBitQ and TurboQuant under a unified comparison framework. We compare the two methods in terms

RIFT: A RubrIc Failure Mode Taxonomy and Automated Diagnostics

ResearchDGX agent

arXiv:2604.01375v2 Announce Type: replace Abstract: Rubric-based evaluation is widely used in LLM benchmarks and training pipelines for open-ended, less verifiable tasks. While prior work has demonstr

Right for the Wrong Reasons: Epistemic Regret Minimization for LLM Causal Reasoning

Model ReleasesDGX agent

arXiv:2602.11675v3 Announce Type: replace Abstract: Large language models may answer causal questions correctly for the wrong reasons, substituting associational shortcuts P(Y|X) for the interventiona

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation

Model ReleasesDGX agent

arXiv:2604.19092v1 Announce Type: cross Abstract: Recent advances in large-scale video world models have enabled increasingly realistic future prediction, raising the prospect of leveraging imagined v

RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian

ApplicationsDGX agent

arXiv:2604.19593v1 Announce Type: cross Abstract: The importance of clear and correct text in legal documents cannot be understated, and, consequently, a grammatical error correction tool meant to ass

S2MAM: Semi-supervised Meta Additive Model for Robust Estimation and Variable Selection

ApplicationsDGX agent

arXiv:2604.19072v1 Announce Type: cross Abstract: Semi-supervised learning with manifold regularization is a classical framework for jointly learning from both labeled and unlabeled data, where the ke

Safety-Critical Contextual Control via Online Riemannian Optimization with World Models

SafetyDGX agent

arXiv:2604.19639v1 Announce Type: cross Abstract: Modern world models are becoming too complex to admit explicit dynamical descriptions. We study safety-critical contextual control, where a Planner mu

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2604.19638v1 Announce Type: new Abstract: Multimodal Large Language Models are increasingly adopted as autonomous agents in interactive environments, yet their ability to proactively address saf

SAGE-32B: Agentic Reasoning via Iterative Distillation

Model ReleasesDGX agent

arXiv:2601.04237v2 Announce Type: replace Abstract: We demonstrate SAGE-32B, a 32 billion parameter language model that focuses on agentic reasoning and long range planning tasks. Unlike chat models t

SAHM: A Benchmark for Arabic Financial and Shari'ah-Compliant Reasoning

Model ReleasesDGX agent

arXiv:2604.19098v1 Announce Type: cross Abstract: English financial NLP has progressed rapidly through benchmarks for sentiment, document understanding, and financial question answering, while Arabic

SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning

Model ReleasesDGX agent

arXiv:2604.19048v1 Announce Type: cross Abstract: The combination of Mixture-of-Experts (MoE) and Low-Rank Adaptation (LoRA) has shown significant potential for enhancing the multi-task learning capab

SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution

Model ReleasesDGX agent

arXiv:2604.18982v1 Announce Type: new Abstract: Social intelligence, the ability to navigate complex interpersonal interactions, presents a fundamental challenge for language agents. Training such age

SCURank: Ranking Multiple Candidate Summaries with Summary Content Units for Enhanced Summarization

ResearchDGX agent

arXiv:2604.19185v1 Announce Type: cross Abstract: Small language models (SLMs), such as BART, can achieve summarization performance comparable to large language models (LLMs) via distillation. However

SEAT: Sparse Entity-Aware Tuning for Knowledge Adaptation while Preserving Epistemic Abstention

SafetyDGX agent

arXiv:2506.14387v3 Announce Type: replace Abstract: Adapting LLMs with new knowledge is increasingly important, but standard fine-tuning often erodes aligned epistemic abstention: the ability to ackno

See2Refine: Vision-Language Feedback Improves LLM-Based eHMI Action Designers

ResearchDGX agent

arXiv:2602.02063v2 Announce Type: replace-cross Abstract: Automated vehicles lack natural communication channels with other road users, making external Human-Machine Interfaces (eHMIs) essential for c

Self-Improving Tabular Language Models via Iterative Group Alignment

Local AiDGX agent

arXiv:2604.18966v1 Announce Type: cross Abstract: While language models have been adapted for tabular data generation, two fundamental limitations remain: (1) static fine-tuning produces models that c

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

SafetyDGX agent

arXiv:2604.18835v1 Announce Type: cross Abstract: We propose a scalable, multifactorial experimental framework that systematically probes LLM sensitivity to subtle semantic changes in pairwise documen

Sentipolis: Emotion-Aware Agents for Social Simulations

ResearchDGX agent

arXiv:2601.18027v2 Announce Type: replace Abstract: LLM agents are increasingly used for social simulation, yet emotion is often treated as a transient cue, causing emotional amnesia and weak long-hor

ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.19254v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) reduces the training cost of full-parameter fine-tuning for large language models (LLMs) by training only a sma

Sherpa.ai Privacy-Preserving Multi-Party Entity Alignment without Intersection Disclosure for Noisy Identifiers

SafetyDGX agent

arXiv:2604.19219v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training among multiple parties without centralizing raw data. There are two main paradigms in FL:

SimDiff: Depth Pruning via Similarity and Difference

ResearchDGX agent

arXiv:2604.19520v1 Announce Type: new Abstract: Depth pruning improves the deployment efficiency of large language models (LLMs) by identifying and removing redundant layers. A widely accepted standar

Skillful Global Ocean Emulation and the Role of Correlation-Aware Loss

ResearchDGX agent

arXiv:2604.18727v1 Announce Type: cross Abstract: Machine learning emulators have shown extraordinary skill in forecasting atmospheric states, and their application to global ocean dynamics offers sim

SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion

Model ReleasesDGX agent

arXiv:2510.17925v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at code-related tasks but often struggle in realistic software repositories, where project-specific APIs an

Speculative End-Turn Detector for Efficient Speech Chatbot Assistant

Local AiDGX agent

arXiv:2503.23439v2 Announce Type: replace-cross Abstract: Spoken dialogue systems powered by large language models have demonstrated remarkable abilities in understanding human speech and generating a

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

HardwareDGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

SPRITE: From Static Mockups to Engine-Ready Game UI

Model ReleasesDGX agent

arXiv:2604.18591v1 Announce Type: cross Abstract: Game UI implementation requires translating stylized mockups into interactive engine entities. However, current 'Screenshot-to-Code' tools often strug

ST-Prune: Training-Free Spatio-Temporal Token Pruning for Vision-Language Models in Autonomous Driving

AgentsDGX agent

arXiv:2604.19145v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become central to autonomous driving systems, yet their deployment is severely bottlenecked by the massive computat

StepFly: Agentic Troubleshooting Guide Automation for Incident Diagnosis

Model ReleasesDGX agent

arXiv:2510.10074v2 Announce Type: replace Abstract: Effective incident management in large-scale IT systems relies on troubleshooting guides (TSGs), but their manual execution is slow and error-prone.

Streamliners for Answer Set Programming

ResearchDGX agent

arXiv:2604.19251v1 Announce Type: cross Abstract: Streamliner constraints reduce the search space of combinatorial problems by ruling out portions of the solution space. We adapt the StreamLLM approac

TACENR: Task-Agnostic Contrastive Explanations for Node Representations

TutorialsDGX agent

arXiv:2604.19372v1 Announce Type: cross Abstract: Graph representation learning has achieved notable success in encoding graph-structured data into latent vector spaces, enabling a wide range of downs

Tadabur: A Large-Scale Quran Audio Dataset

ResearchDGX agent

arXiv:2604.18932v1 Announce Type: cross Abstract: Despite growing interest in Quranic data research, existing Quran datasets remain limited in both scale and diversity. To address this gap, we present

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

Model ReleasesDGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

Model ReleasesDGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.18296v2 Announce Type: replace-cross Abstract: Temporal Knowledge Graph Question Answering (TKGQA) is inherently challenging, as it requires sophisticated reasoning over dynamic facts with

Temporal UI State Inconsistency in Desktop GUI Agents: Formalizing and Defending Against TOCTOU Attacks on Computer-Use Agents

ResearchDGX agent

arXiv:2604.18860v1 Announce Type: cross Abstract: GUI agents that control desktop computers via screenshot-and-click loops introduce a new class of vulnerability: the observation-to-action gap (mean 6

Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters

HardwareDGX agent

arXiv:2509.18831v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have significantly improved image and video synthesis. In addition, several concept control methods have b

TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction

AgentsDGX agent

arXiv:2602.06400v2 Announce Type: replace-cross Abstract: The prediction of 3D semantic occupancy enables autonomous vehicles (AVs) to perceive the fine-grained geometric and semantic scene structure

The Cost of Relaxation: Evaluating the Error in Convex Neural Network Verification

ResearchDGX agent

arXiv:2604.18728v1 Announce Type: cross Abstract: Many neural network (NN) verification systems represent the network's input-output relation as a constraint program. Sound and complete, representatio

The data heat island effect: quantifying the impact of AI data centers in a warming world

Local AiDGX agent

arXiv:2603.20897v3 Announce Type: replace-cross Abstract: The strong and continuous increase of AI-based services leads to the steady proliferation of AI data centres worldwide with the unavoidable es

The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models

Model ReleasesDGX agent

arXiv:2604.19139v1 Announce Type: cross Abstract: As Large Language Models (LLMs) continue to evolve through alignment techniques such as Reinforcement Learning from Human Feedback (RLHF) and Constitu

The Triadic Loop: A Framework for Negotiating Alignment in AI Co-hosted Livestreaming

SafetyDGX agent

arXiv:2604.18850v1 Announce Type: cross Abstract: AI systems are increasingly embedded in multi-user social environments, yet most alignment frameworks conceptualize interaction as a dyadic relationsh

Thermal Anomaly Detection using Physics Aware Neuromorphic Networks: Comparison between Raw and L1C Sentinel-2 Data

ResearchDGX agent

arXiv:2604.18606v1 Announce Type: cross Abstract: Damage caused by bushfires and volcanic eruptions escalates rapidly when detection is delayed, making fast and reliable early warning capabilities ess

Think Before Writing: Feature-Level Multi-Objective Optimization for Generative Citation Visibility

ResearchDGX agent

arXiv:2604.19113v1 Announce Type: cross Abstract: Generative answer engines expose content through selective citation rather than ranked retrieval, fundamentally altering how visibility is determined.

Time Series Augmented Generation for Financial Applications

Model ReleasesDGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

Towards Auto-Building of Embedded FPGA-based Soft Sensors for Wastewater Flow Estimation

Local AiDGX agent

arXiv:2407.05102v2 Announce Type: replace-cross Abstract: Executing flow estimation using Deep Learning (DL)-based soft sensors on resource-limited IoT devices has demonstrated promise in terms of rel

Towards Energy Impact on AI-Powered 6G IoT Networks: Centralized vs. Decentralized

ApplicationsDGX agent

arXiv:2604.19377v1 Announce Type: new Abstract: The emergence of sixth-generation (6G) technologies has introduced new challenges and opportunities for machine learning (ML) applications in Internet o

Towards Generalization of Graph Neural Networks for AC Optimal Power Flow

ResearchDGX agent

arXiv:2510.06860v2 Announce Type: replace-cross Abstract: AC Optimal Power Flow (ACOPF) is computationally intensive for large-scale grids, often requiring prohibitive solution times with conventional

Towards Optimal Agentic Architectures for Offensive Security Tasks

Model ReleasesDGX agent

arXiv:2604.18718v1 Announce Type: cross Abstract: Agentic security systems increasingly audit live targets with tool-using LLMs, but prior systems fix a single coordination topology, leaving unclear w

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

Model ReleasesDGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model

ResearchDGX agent

arXiv:2604.19635v1 Announce Type: cross Abstract: While generative models have set new benchmarks for Target Speaker Extraction (TSE), their inherent reliance on global context precludes deployment in

← Previous
1…319320321322323…354
Next →