AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Research

Assessing Dutch Syllabification Algorithms and Improving Accuracy by Combining Phonetic and Orthographic Information through Deep Learning

DGX agent

arXiv:2605.28834v1 Announce Type: cross Abstract: Syllabification describes the task of dividing words into syllables. Due to many rules and exceptions, training an algorithm to perform syllabificatio

researcharxiv-cs-ai
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Crystalline Materials

DGX agent

arXiv:2510.04704v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promising potential in scientific research, enabling tasks ranging from knowledge retrieval to propert

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence

DGX agent

arXiv:2605.21739v2 Announce Type: replace Abstract: Emotional intelligence (EI), the ability to perceive, understand, and respond appropriately to others' emotional states, is central to human communi

model-releasesarxiv-cs-ai
29 May 2026
Safety

Audio Jailbreaks in Large Audio-Language Models: Taxonomy, Attack-Defense Analysis, and Cost-Aware Evaluation

DGX agent

arXiv:2605.30031v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) expand jailbreak risks from token-level prompting to the full speech perception-to-reasoning pipeline, where unsaf

safetyarxiv-cs-ai
29 May 2026
Safety

Automating Low-Risk Code Review at Meta: RADAR, Risk Calibration, and Review Efficiency

DGX agent

arXiv:2605.30208v1 Announce Type: cross Abstract: AI-assisted coding tools have altered software production. At Meta, significant lines of code per human-landed diff grew by 105.9% year over year and

safetyarxiv-cs-ai
29 May 2026
Model Releases

AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents

DGX agent

arXiv:2602.02849v2 Announce Type: replace Abstract: The design of Analog and Mixed-Signal (AMS) integrated circuits remains heavily reliant on expert knowledge, with transistor sizing a major bottlene

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Balancing Multimodal Learning through Label Space Reshaping

DGX agent

arXiv:2605.28869v1 Announce Type: cross Abstract: Multimodal learning often suffers from modality imbalance, where modalities that converge faster dominate optimization while others remain undertraine

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

DGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

model-releasesarxiv-cs-ai
29 May 2026
Safety

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

DGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

safetyarxiv-cs-ai
29 May 2026
Applications

Before the Shutter: Aesthetic and Actionable Portrait Photography Planning in 3D Scenes

DGX agent

arXiv:2605.30318v1 Announce Type: cross Abstract: Portrait photography is largely decided before the shutter opens: the subject's pose, the camera configuration, and the lighting devices must be coord

applicationsarxiv-cs-ai
29 May 2026
Local Ai

Behavior-Aware Auxiliary Corrections for Off-Policy Temporal-Difference Prediction

DGX agent

arXiv:2605.28855v1 Announce Type: new Abstract: Temporal-difference learning with function approximation can be unstable under off-policy sampling. TDC stabilizes off-policy TD through an auxiliary co

local-aiarxiv-cs-ai
29 May 2026
Safety

Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction

DGX agent

arXiv:2605.28849v1 Announce Type: new Abstract: Gradient temporal-difference methods provide stable off-policy prediction with linear function approximation, but their practical performance is strongl

safetyarxiv-cs-ai
29 May 2026
Local Ai

Benchmarking at the Edge of Comprehension

DGX agent

arXiv:2602.14307v3 Announce Type: replace Abstract: As frontier Large Language Models (LLMs) increasingly saturate new benchmarks shortly after they are published, benchmarking itself is at a juncture

local-aiarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset

DGX agent

arXiv:2605.29462v1 Announce Type: cross Abstract: The emergence of Large Vision-Language Models (LVLMs) has substantially expanded model capabilities beyond text-only understanding, enabling unified i

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

DGX agent

arXiv:2509.23571v3 Announce Type: replace-cross Abstract: As cyber threats continue to grow in scale and sophistication, blue team defenders increasingly require advanced tools to proactively detect a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation

DGX agent

arXiv:2605.28830v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in safety-critical applications, robust content moderation becomes essential. We present a c

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Positional Encoding Strategies for Transformer-Based EEG Foundation Models

DGX agent

arXiv:2605.29754v1 Announce Type: new Abstract: Electroencephalography (EEG) is a widely used non-invasive technique for measuring brain activity in brain-computer interface (BCI) applications. Superv

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

DGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

model-releasesarxiv-cs-ai
29 May 2026
Research

Better Later Than Sooner: Neuro-Symbolic Knowledge Graph Construction via Ontology-grounded Post-extraction Correction

DGX agent

arXiv:2605.29168v1 Announce Type: new Abstract: Question answering (QA) is a core challenge in AI, particularly for complex queries requiring multi-hop reasoning across documents, or symbolic operatio

researcharxiv-cs-ai
29 May 2026
Research

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning

DGX agent

arXiv:2605.30231v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) often struggle with robust 3D spatial reasoning. Prevailing methods that rely on fine-tuning with 3D visual question-ans

researcharxiv-cs-ai
29 May 2026
Research

Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?

DGX agent

arXiv:2510.16060v2 Announce Type: replace-cross Abstract: The recent development of foundation models for time series data has generated considerable interest in using such models across a variety of

researcharxiv-cs-ai
29 May 2026
Safety

Beyond Attack Success Rate: Temporal Logit Observability for LLM Safety Failures

DGX agent

arXiv:2605.29629v1 Announce Type: new Abstract: Attack Success Rate (ASR) evaluates each jailbreak with a single yes/no label at the end of generation, telling us whether a failure happened but not ho

safetyarxiv-cs-ai
29 May 2026
Safety

Beyond Bilingual Transfer: Multilingual Code-Switching in Instruction Tuning

DGX agent

arXiv:2605.29414v1 Announce Type: cross Abstract: Recent studies have shown that code-switching data (CSD), in which multiple languages are mixed within the same context, can improve cross-lingual tra

safetyarxiv-cs-ai
29 May 2026
Agents

Beyond Consensus: Trace-Level Synthesis in Mixture of Agents

DGX agent

arXiv:2605.29116v1 Announce Type: new Abstract: When multiple LLM agents solve the same problem, standard practice compresses each agent's reasoning into a majority vote or layered synthesis, treating

agentsarxiv-cs-ai
29 May 2026
Research

Beyond MSE: Improving Precipitation Nowcasting with Multi-Quantile Regression

DGX agent

arXiv:2605.30122v1 Announce Type: cross Abstract: Deep-learning precipitation nowcasting models are often optimized using pointwise losses such as mean squared error or mean absolute error, which can

researcharxiv-cs-ai
29 May 2026
Research

Beyond Normalization: Rethinking the Partition Function as a Difficulty Scheduler for RLVR

DGX agent

arXiv:2602.12642v2 Announce Type: replace-cross Abstract: Reward-maximizing RL methods have shown to be capable of enhancing the reasoning performance of LLMs, but often lead to reduced generation div

researcharxiv-cs-ai
29 May 2026
Model Releases

Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization

DGX agent

arXiv:2605.28969v1 Announce Type: cross Abstract: If an AI agent makes decisions on a person's behalf, those decisions must align with its user. We introduce representational accuracy to measure how f

model-releasesarxiv-cs-ai
29 May 2026
Safety

Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling

DGX agent

arXiv:2605.29697v1 Announce Type: new Abstract: In Agentic Search, trajectory-level outcome rewards fail to quantify the behavioral contributions of individual steps, while existing step-level reward

safetyarxiv-cs-ai
29 May 2026
Tutorials

BioArc: Discovering Optimal Neural Architectures for Biological Foundation Models

DGX agent

arXiv:2512.00283v3 Announce Type: replace-cross Abstract: Foundation models have revolutionized various fields such as natural language processing (NLP) and computer vision (CV). While efforts have be

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

BioRefusalAudit: Auditing Biosecurity Refusal Depth Using General and Domain-Fine-Tuned Sparse Autoencoders

DGX agent

arXiv:2605.30162v1 Announce Type: new Abstract: Biosecurity evaluations of language models typically ask whether models produce hazardous output. This paper asks a complementary question: when a model

model-releasesarxiv-cs-ai
29 May 2026
Agents

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

DGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

agentsarxiv-cs-ai
29 May 2026
Local Ai

BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inference

DGX agent

arXiv:2605.29233v1 Announce Type: cross Abstract: Diffusion language models (dLLMs) generate text by iteratively denoising multiple token positions in parallel, offering an attractive alternative to s

local-aiarxiv-cs-ai
29 May 2026
Safety

BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models

DGX agent

arXiv:2605.30226v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for grounding visual-language understanding into real-world robotic manipulat

safetyarxiv-cs-ai
29 May 2026
Model Releases

Brain-IT-VQA: From Brain Signals to Answers

DGX agent

arXiv:2605.29588v1 Announce Type: cross Abstract: Decoding visual content from fMRI signals recorded while a person views images, and specifically answering questions about the seen images, is a long-

model-releasesarxiv-cs-ai
29 May 2026
Research

Bridge-RAG: An Abstract Bridge Tree Based Retrieval Augmented Generation Algorithm

DGX agent

arXiv:2603.26668v2 Announce Type: replace-cross Abstract: As an important paradigm for enhancing the generation quality of Large Language Models (LLMs), retrieval-augmented generation (RAG) faces the

researcharxiv-cs-ai
29 May 2026
Model Releases

Bridging the Semantic Gap for Categorical Data Clustering via Large Language Models

DGX agent

arXiv:2601.01162v3 Announce Type: replace-cross Abstract: Qualitative data are widespread in domains such as healthcare, marketing, and bioinformatics, where clustering offers a fundamental tool for p

model-releasesarxiv-cs-ai
29 May 2026
Safety

Bridging the Sim-to-Real Gap in Reinforcement Learning-Based Industrial Dispatching through Execution Semantics

DGX agent

arXiv:2605.29078v1 Announce Type: new Abstract: Event-driven scheduling policies are increasingly deployed in industrial environments, where decisions are made under asynchronous and partially observe

safetyarxiv-cs-ai
29 May 2026
Hardware

CA-AC-MPC: CUDA-Accelerated Actor-Critic Model Predictive Control

DGX agent

arXiv:2605.29155v1 Announce Type: cross Abstract: In the literature, actor-critic model predictive control (AC-MPC) integrates MPC with reinforcement learning to enable high-performance control of com

hardwarearxiv-cs-ai
29 May 2026
Model Releases

CalArena: A Large-Scale Post-Hoc Calibration Benchmark

DGX agent

arXiv:2605.30188v1 Announce Type: cross Abstract: Reliable probability estimates are critical in many machine learning applications, yet modern classifiers are often poorly calibrated. Post-hoc calibr

model-releasesarxiv-cs-ai
29 May 2026
Safety

Causal-JEPA: Learning World Models through Object-Level Latent Masking

DGX agent

arXiv:2602.11389v2 Announce Type: replace Abstract: World models require robust relational understanding to support prediction, reasoning, and control. While object-centric representations provide a u

safetyarxiv-cs-ai
29 May 2026
Research

Causal Label Recovery in Payment Networks

DGX agent

arXiv:2605.29272v1 Announce Type: cross Abstract: Fraud detection models in payment networks train on chargeback labels that are systematically biased. Every label must survive three sequential gates:

researcharxiv-cs-ai
29 May 2026
Safety

CB-SLICE: Concept-Based Interpretable Error Slice Discovery

DGX agent

arXiv:2605.29836v1 Announce Type: cross Abstract: Despite strong average-case performance, deep learning models often exhibit systematic errors on specific population groups, known as error slices. Id

safetyarxiv-cs-ai
29 May 2026
Safety

Certified Policy Optimisation for Nested Causal Bandits via PAC-Bayes Risk

DGX agent

arXiv:2605.29788v1 Announce Type: new Abstract: Critical sequential decisions are rarely single-timescale: a strategic decision causally shapes the context in which every subsequent tactical choice is

safetyarxiv-cs-ai
29 May 2026
Model Releases

Citation-Closure Retrieval and Per-Rule Attribution for Real-World Regulatory Compliance Question Answering

DGX agent

arXiv:2605.29742v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) for regulatory compliance demands rigorous traceability via comprehensive citations across multi-tiered authority

model-releasesarxiv-cs-ai
29 May 2026
Research

City-Mesh3R: Simulation-Ready City-Scale 3D Mesh Reconstruction from Multi-View Images

DGX agent

arXiv:2605.30310v1 Announce Type: cross Abstract: City-scale 3D surface reconstruction from multiview images for downstream 3D simulation, poses highly challenging problems due to the scale and comple

researcharxiv-cs-ai
29 May 2026
Model Releases

CityGen: Structure-Guided City-Style Synthesis for Cross-City Autonomous Driving

DGX agent

arXiv:2605.29935v1 Announce Type: cross Abstract: Autonomous driving systems are commonly trained and evaluated within limited geographic regions, which hinders their scalability when deployed in new

model-releasesarxiv-cs-ai
29 May 2026
Agents

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

DGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

agentsarxiv-cs-ai
29 May 2026
Model Releases

CodeEvolve: an open source evolutionary coding agent for algorithmic discovery and optimization

DGX agent

arXiv:2510.14150v5 Announce Type: replace Abstract: We introduce CodeEvolve, an open-source framework that couples large language models with island-based evolutionary search for end-to-end algorithmi

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…240241242243244…452
Next →