AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

Assessing Capabilities of Large Language Models in Social Media Analytics: A Multi-task Quest

DGX agent

arXiv:2604.18955v1 Announce Type: cross Abstract: In this study, we present the first comprehensive evaluation of modern LLMs - including GPT-4, GPT-4o, GPT-3.5-Turbo, Gemini 1.5 Pro, DeepSeek-V3, Lla

model-releasesarxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Attention-based Multi-modal Deep Learning Model of Spatio-temporal Crop Yield Prediction with Satellite, Soil and Climate Data

DGX agent

arXiv:2604.19217v1 Announce Type: cross Abstract: Crop yield prediction is one of the most important challenge, which is crucial to world food security and policy-making decisions. The conventional fo

safetyarxiv-cs-ai
22 Apr 2026
Safety

AutoAWG: Adverse Weather Generation with Adaptive Multi-Controls for Automotive Videos

DGX agent

arXiv:2604.18993v1 Announce Type: cross Abstract: Perception robustness under adverse weather remains a critical challenge for autonomous driving, with the core bottleneck being the scarcity of real-w

safetyarxiv-cs-ai
22 Apr 2026
Agents

Autogenesis: A Self-Evolving Agent Protocol

DGX agent

arXiv:2604.15034v2 Announce Type: replace Abstract: Recent advances in LLM based agent systems have shown promise in tackling complex, long horizon tasks. However, existing agent protocols (e.g., A2A

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

AutomationBench

DGX agent

arXiv:2604.18934v1 Announce Type: new Abstract: Existing AI benchmarks for software automation rarely combine cross-application coordination, autonomous API discovery, and policy adherence. Real busin

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

BAPO: Boundary-Aware Policy Optimization for Reliable Agentic Search

DGX agent

arXiv:2601.11037v2 Announce Type: replace Abstract: RL-based agentic search enables LLMs to solve complex questions via dynamic planning and external search. While this approach significantly enhances

safetyarxiv-cs-ai
22 Apr 2026
Research

BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps

DGX agent

arXiv:2604.19532v1 Announce Type: cross Abstract: Tokenizing music to fit the general framework of language models is a compelling challenge, especially considering the diverse symbolic structures in

researcharxiv-cs-ai
22 Apr 2026
Research

BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design

DGX agent

arXiv:2508.21184v3 Announce Type: replace-cross Abstract: We propose a general-purpose approach for improving the ability of large language models (LLMs) to intelligently and adaptively gather informa

researcharxiv-cs-ai
22 Apr 2026
Safety

Benchmarking Misuse Mitigation Against Covert Adversaries

DGX agent

arXiv:2506.06414v2 Announce Type: replace-cross Abstract: Existing language model safety evaluations focus on overt attacks and low-stakes tasks. In reality, an attacker can easily subvert existing sa

safetyarxiv-cs-ai
22 Apr 2026
Applications

Benign Overfitting in Adversarial Training for Vision Transformers

DGX agent

arXiv:2604.19724v1 Announce Type: cross Abstract: Despite the remarkable success of Vision Transformers (ViTs) across a wide range of vision tasks, recent studies have revealed that they remain vulner

applicationsarxiv-cs-ai
22 Apr 2026
Agents

Best Agent Identification for General Game Playing

DGX agent

arXiv:2507.00451v2 Announce Type: replace-cross Abstract: We present an efficient and generalised procedure to accurately identify the best (or near best) performing algorithm for each sub-task in a m

agentsarxiv-cs-ai
22 Apr 2026
Applications

Beyond Coefficients: Forecast-Necessity Testing for Interpretable Causal Discovery in Nonlinear Time-Series Models

DGX agent

arXiv:2604.18751v1 Announce Type: cross Abstract: Nonlinear machine-learning models are increasingly used to discover causal relationships in time-series data, yet the interpretation of their outputs

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

Beyond Explicit Refusals: Soft-Failure Attacks on Retrieval-Augmented Generation

DGX agent

arXiv:2604.18663v1 Announce Type: cross Abstract: Existing jamming attacks on Retrieval-Augmented Generation (RAG) systems typically induce explicit refusals or denial-of-service behaviors, which are

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Beyond Itinerary Planning-A Real-World Benchmark for Multi-Turn and Tool-Using Travel Tasks

DGX agent

arXiv:2512.22673v3 Announce Type: replace Abstract: Travel planning is a natural real-world task to test large language models' (LLMs) planning and tool-use abilities. Although prior work has studied

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Beyond One Output: Visualizing and Comparing Distributions of Language Model Generations

DGX agent

arXiv:2604.18724v1 Announce Type: new Abstract: Users typically interact with and evaluate language models via single outputs, but each output is just one sample from a broad distribution of possible

researcharxiv-cs-ai
22 Apr 2026
Safety

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

DGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

safetyarxiv-cs-ai
22 Apr 2026
Safety

Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration

DGX agent

arXiv:2604.17457v2 Announce Type: replace-cross Abstract: Dynamic programming is one of the most fundamental methodologies for solving Markov decision problems. Among its many variants, Q-value iterat

safetyarxiv-cs-ai
22 Apr 2026
Agents

Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development

DGX agent

arXiv:2604.16323v2 Announce Type: replace-cross Abstract: As autonomous coding agents become deeply embedded in software development workflows, their high operational velocity introduces a critical ov

agentsarxiv-cs-ai
22 Apr 2026
Research

Bootstrapping Code Translation with Weighted Multilanguage Exploration

DGX agent

arXiv:2601.03512v2 Announce Type: replace-cross Abstract: Code translation across multiple programming languages is essential yet challenging due to two vital obstacles: scarcity of parallel data pair

researcharxiv-cs-ai
22 Apr 2026
Applications

Bridging the High-Frequency Data Gap: A Millisecond-Resolution Network Dataset for Advancing Time Series Foundation Models

DGX agent

arXiv:2603.16497v2 Announce Type: replace-cross Abstract: Time series foundation models (TSFMs) require diverse, real-world datasets to adapt across varying domains and temporal frequencies. However,

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmark

DGX agent

arXiv:2505.16968v4 Announce Type: replace-cross Abstract: Cross-architecture GPU code transpilation is essential for unlocking low-level hardware portability, yet no scalable solution exists. We intro

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

CentaurTA Studio: A Self-Improving Human-Agent Collaboration System for Thematic Analysis

DGX agent

arXiv:2604.18589v1 Announce Type: cross Abstract: Thematic analysis is difficult to scale: manual workflows are labor-intensive, while fully automated pipelines often lack controllability and transpar

safetyarxiv-cs-ai
22 Apr 2026
Safety

Chain-of-Thought as a Lens: Evaluating Structured Reasoning Alignment between Human Preferences and Large Language Models

DGX agent

arXiv:2511.06168v3 Announce Type: replace Abstract: This paper primarily demonstrates a method to quantitatively assess the alignment between multi-step, structured reasoning in large language models

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Characterizing AlphaEarth Embedding Geometry for Agentic Environmental Reasoning

DGX agent

arXiv:2604.18715v1 Announce Type: cross Abstract: Earth observation foundation models encode land surface information into dense embedding vectors, yet the geometric structure of these representations

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Language

DGX agent

arXiv:2604.19667v1 Announce Type: cross Abstract: At present, executable visual workflows have emerged as a mainstream paradigm in real-world industrial deployments, offering strong reliability and co

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence

DGX agent

arXiv:2602.12851v3 Announce Type: replace-cross Abstract: Deploying expressive learning models directly on programmable dataplanes promises line-rate, low-latency traffic analysis but remains hindered

researcharxiv-cs-ai
22 Apr 2026
Research

Choose Your Own Adventure: Non-Linear AI-Assisted Programming with EvoGraph

DGX agent

arXiv:2604.18883v1 Announce Type: cross Abstract: Current AI-assisted programming tools are predominantly linear and chat-based, which deviates from the iterative and branching nature of programming i

researcharxiv-cs-ai
22 Apr 2026
Agents

ClawNet: Human-Symbiotic Agent Network for Cross-User Autonomous Cooperation

DGX agent

arXiv:2604.19211v1 Announce Type: new Abstract: Current AI agent frameworks have made remarkable progress in automating individual tasks, yet all existing systems serve a single user. Human productivi

agentsarxiv-cs-ai
22 Apr 2026
Safety

Cloning Deterministic Worlds: The Critical Role of Latent Geometry in Long-Horizon World Models

DGX agent

arXiv:2510.26782v3 Announce Type: replace-cross Abstract: A world model is an internal model that simulates how the world evolves. Given past observations and actions, it predicts the future physical

safetyarxiv-cs-ai
22 Apr 2026
Research

Co-Refine: AI-Powered Tool Supporting Qualitative Analysis

DGX agent

arXiv:2604.19309v1 Announce Type: cross Abstract: Qualitative coding relies on a researcher's application of codes to textual data. As coding proceeds across large datasets, interpretations of codes o

researcharxiv-cs-ai
22 Apr 2026
Research

CoCo-SAM3: Harnessing Concept Conflict in Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2604.19648v1 Announce Type: cross Abstract: SAM3 advances open-vocabulary semantic segmentation by introducing a prompt-driven mask generation paradigm. However, in multi-class open-vocabulary s

researcharxiv-cs-ai
22 Apr 2026
Applications

CoDA: Towards Effective Cross-domain Knowledge Transfer via CoT-guided Domain Adaptation

DGX agent

arXiv:2604.19488v1 Announce Type: new Abstract: Large language models (LLMs) have achieved substantial advances in logical reasoning, yet they continue to lag behind human-level performance. In-contex

applicationsarxiv-cs-ai
22 Apr 2026
Local Ai

COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition

DGX agent

arXiv:2503.07259v2 Announce Type: replace-cross Abstract: The goal of creating intelligent, human-centered wearable systems for continuous activity understanding faces a fundamental trade-off: Egocent

local-aiarxiv-cs-ai
22 Apr 2026
Model Releases

Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs

DGX agent

arXiv:2604.18587v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated significant potential in formal theorem proving, yet state-of-the-art performance often necessitates pr

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Conjuring Semantic Similarity

DGX agent

arXiv:2410.16431v4 Announce Type: replace Abstract: The semantic similarity between sample expressions measures the distance between their latent 'meaning'. These meanings are themselves typically rep

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Council Mode: Mitigating Hallucination and Bias in LLMs via Multi-Agent Consensus

DGX agent

arXiv:2604.02923v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly those employing Mixture-of-Experts (MoE) architectures, have achieved remarkable capabilities acros

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

CounterRefine: Answer-Conditioned Counterevidence Retrieval for Inference-Time Knowledge Repair in Factual Question Answering

DGX agent

arXiv:2603.16091v2 Announce Type: replace-cross Abstract: In factual question answering, many errors are not failures of access but failures of commitment: the system retrieves relevant evidence, yet

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Counting Worlds Branching Time Semantics for post-hoc Bias Mitigation in generative AI

DGX agent

arXiv:2604.19431v1 Announce Type: cross Abstract: Generative AI systems are known to amplify biases present in their training data. While several inference-time mitigation strategies have been propose

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Cross-Model Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Across Three Large Language Models

DGX agent

arXiv:2604.19598v1 Announce Type: cross Abstract: This study compared repeated generation consistency of exercise prescription outputs across three large language models (LLMs), specifically GPT-4.1,

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

DGX agent

arXiv:2604.19262v1 Announce Type: cross Abstract: Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities.

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Curiosity-Critic: Cumulative Prediction Error Improvement as a Tractable Intrinsic Reward for World Model Training

DGX agent

arXiv:2604.18701v1 Announce Type: cross Abstract: Local prediction-error-based curiosity rewards focus on the current transition without considering the world model's cumulative prediction error acros

local-aiarxiv-cs-ai
22 Apr 2026
Safety

Curvature-Aware PCA with Geodesic Tangent Space Aggregation for Semi-Supervised Learning

DGX agent

arXiv:2604.18816v1 Announce Type: cross Abstract: Principal Component Analysis (PCA) is a fundamental tool for representation learning, but its global linear formulation fails to capture the structure

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps

DGX agent

arXiv:2604.19533v1 Announce Type: cross Abstract: We introduce the Cyber Defense Benchmark, a benchmark for measuring how well large language model (LLM) agents perform the core SOC analyst task of th

model-releasesarxiv-cs-ai
22 Apr 2026
Research

DanceCrafter: Fine-Grained Text-Driven Controllable Dance Generation via Choreographic Syntax

DGX agent

arXiv:2604.18648v1 Announce Type: cross Abstract: Text-driven controllable dance generation remains under-explored, primarily due to the severe scarcity of high-quality datasets and the inherent diffi

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?

DGX agent

arXiv:2602.18571v2 Announce Type: replace-cross Abstract: While significant progress has been made in automating various aspects of software development through coding agents, there is still significa

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Decompose, Structure, and Repair: A Neuro-Symbolic Framework for Autoformalization via Operator Trees

DGX agent

arXiv:2604.19000v1 Announce Type: cross Abstract: Statement autoformalization acts as a critical bridge between human mathematics and formal mathematics by translating natural language problems into f

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs

DGX agent

arXiv:2511.22099v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained sett

safetyarxiv-cs-ai
22 Apr 2026
Research

Design Rules for Extreme-Edge Scientific Computing on AI Engines

DGX agent

arXiv:2604.19106v1 Announce Type: cross Abstract: Extreme-edge scientific applications use machine learning models to analyze sensor data and make real-time decisions. Their stringent latency and thro

researcharxiv-cs-ai
22 Apr 2026
← Previous
1…395396397398399…443
Next →