AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

Steganography Without Modification: Hidden Communication via LLM Seeds

DGX agent

arXiv:2606.09135v1 Announce Type: cross Abstract: We demonstrate that widely deployed Large Language Model (LLM) inference stacks harbor a steganographic channel that requires no modification to model

safetyarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

STELLAR: Spatio-Temporal Environmental Learning with Latent Alignment and Refinement for Long-Tailed Species Distribution Modeling

DGX agent

arXiv:2606.08484v1 Announce Type: cross Abstract: Joint Species Distribution Modeling (JSDM) is a key enabler for biodiversity monitoring and conservation planning. However, accurate JSDM faces two co

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

DGX agent

arXiv:2606.07889v1 Announce Type: cross Abstract: LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a saf

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

Strategic Integration of Artificial Intelligence in the C-Suite: The Role of the Chief AI Officer

DGX agent

arXiv:2407.10247v3 Announce Type: replace-cross Abstract: The integration of Artificial Intelligence (AI) into corporate strategy has become critical for organizations seeking to maintain competitive

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

Stress-testing medical large language models reveals latent safety pathology beyond benchmark accuracy

DGX agent

arXiv:2606.07929v1 Announce Type: new Abstract: Large language models (LLMs) are entering clinical practice based on benchmark accuracy that may fail to detect safety-relevant failure modes. Here we p

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning

DGX agent

arXiv:2606.08735v1 Announce Type: new Abstract: Quality-diversity reinforcement learning (QD-RL) aims to construct policy repertoires that contain both high-performing and behaviorally diverse policie

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Structured Neuron Pruning in Deep Neural Networks Using Multi-Armed Bandits

DGX agent

arXiv:2606.07615v1 Announce Type: cross Abstract: Deep neural networks often contain redundant hidden units. Removing individual weights can reduce parameter count, but unstructured sparsity is not al

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Structuring agentic AI for HPC code modernization

DGX agent

arXiv:2606.08710v1 Announce Type: cross Abstract: Modernization of legacy scientific codes is often necessary to keep up with the ever-evolving changes in the compute resource ecosystem. Parallelizati

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)

DGX agent

arXiv:2606.07608v1 Announce Type: cross Abstract: We present a systematic study of fine-tuning OpenAI's Whisper large-v3 for Swiss German ASR, using 1,367 hours of broadcast speech paired with Standar

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Summarization is Not Dead Yet

DGX agent

arXiv:2606.08000v1 Announce Type: cross Abstract: The progress of large language models (LLMs) has fueled claims that model-generated summaries rival or even surpass human-written references, raising

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Supracompetitive Pricing Under AI Monoculture

DGX agent

arXiv:2601.01279v3 Announce Type: replace-cross Abstract: When competing sellers delegate pricing to a shared AI model, such as a large language model, correlated recommendations combined with perform

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SurfDesign: Effective Protein Design on Molecular Surfaces

DGX agent

arXiv:2606.07567v1 Announce Type: cross Abstract: Protein function is largely determined by molecular surface geometry and physicochemical complementarity, yet most protein design methods condition on

model-releasesarxiv-cs-ai
9 Jun 2026
Tutorials

Sustainability and Artificial Intelligence: Necessary, Challenging, and Promising Intersections

DGX agent

arXiv:2606.09006v1 Announce Type: cross Abstract: Both digital economy and digital technology researchers increasingly recognize the need to better address the role that artificial intelligence (AI) p

tutorialsarxiv-cs-ai
9 Jun 2026
Research

SVRG and Beyond via Posterior Correction

DGX agent

arXiv:2512.01930v2 Announce Type: replace-cross Abstract: Stochastic Variance Reduced Gradient (SVRG) and its variants aim to speed-up training by using gradient corrections. Originally proposed over

researcharxiv-cs-ai
9 Jun 2026
Model Releases

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

DGX agent

arXiv:2606.07682v1 Announce Type: cross Abstract: AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex env

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Topics, and Models

DGX agent

arXiv:2606.08451v1 Announce Type: cross Abstract: Safety-aligned large language models often exhibit sycophancy, which is the tendency to affirm users' opinions regardless of factual accuracy. Althoug

safetyarxiv-cs-ai
9 Jun 2026
Agents

Syll: Open-Source Personal Automation with Cross-Surface Execution

DGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

agentsarxiv-cs-ai
9 Jun 2026
Safety

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

DGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

DGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

DGX agent

arXiv:2606.09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

DGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks

DGX agent

arXiv:2510.16028v4 Announce Type: replace-cross Abstract: Neural networks increasingly run on hardware outside the user's control (cloud GPUs, inference marketplaces). Yet ML-as-a-Service reveals litt

hardwarearxiv-cs-ai
9 Jun 2026
Safety

Targeting World Models to Compromise Robot Learning Pipelines

DGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

safetyarxiv-cs-ai
9 Jun 2026
Research

TeamHerald@CHIPSAL 2026: Hate Speech Detection and Sentiment Analysis of Nepali Memes using Transformer-based Architectures and Ensemble Learning

DGX agent

arXiv:2606.08770v1 Announce Type: cross Abstract: The analysis of internet memes in the Nepali language is complicated by frequent code-mixing and a lack of established baseline resources. While memes

researcharxiv-cs-ai
9 Jun 2026
Model Releases

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

DGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Test-Time Adaptive Composition for Machine Learning as a Service (MLaaS) in IoT Environments

DGX agent

arXiv:2606.07685v1 Announce Type: cross Abstract: The dynamic nature of Internet of Things (IoT) environments affects the long-term effectiveness of Machine Learning as a Service (MLaaS) compositions.

researcharxiv-cs-ai
9 Jun 2026
Safety

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

DGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

safetyarxiv-cs-ai
9 Jun 2026
Safety

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

DGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

DGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

model-releasesarxiv-cs-ai
9 Jun 2026
Tutorials

The CIFAR Synthetic Evidence Corpus for Detecting AI-Generated Evidence

DGX agent

arXiv:2606.07916v1 Announce Type: new Abstract: The growing ability of generative models to produce realistic documents poses a direct challenge to evidentiary workflows in the justice system and the

tutorialsarxiv-cs-ai
9 Jun 2026
Safety

The Confidence Trap: Calibration Attacks for Graph Neural Networks

DGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Cross-Architecture Substrate: A Domain-Transcendent, Calibration-Surviving Geometric Invariant of Modern Vision Encoders

DGX agent

arXiv:2606.07882v1 Announce Type: cross Abstract: Different vision neural networks -- trained to classify, contrast, reconstruct, or match images to text -- should have correspondingly different inter

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

DGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective Default Lock-In

DGX agent

arXiv:2606.08172v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate high-stakes interactions in finance, medicine, and mental-health support, yet users have limited con

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models

DGX agent

arXiv:2606.07861v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) excel at multimodal understanding and reasoning, yet their fine-grained visual perception remains underexplored.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

The Montparnasse Algorithm for RNA Design

DGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

DGX agent

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

agentsarxiv-cs-ai
9 Jun 2026
Research

The Topological Dual of a Dataset: A Logic-to-Topology Encoding for AlphaGeometry-Style Data

DGX agent

arXiv:2604.18050v2 Announce Type: replace Abstract: AlphaGeometry represents a milestone in neuro-symbolic reasoning, yet its architecture faces a log-linear scaling bottleneck within its symbolic ded

researcharxiv-cs-ai
9 Jun 2026
Model Releases

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

DGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Think Before You Act: Intention-Guided Reasoning for LLM-Based Location Prediction

DGX agent

arXiv:2606.08122v1 Announce Type: new Abstract: Predicting a user's next Point-of-Interest (POI) based on their historical check-in records is a fundamental task in location-based services. While rece

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning

DGX agent

arXiv:2601.04805v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have attracted much attention due to their exceptional performance. However, their performance mainly stems from think

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

TianJi-Environ: An Autonomous AI Scientist for Atmospheric Environmental Research

DGX agent

arXiv:2606.07697v1 Announce Type: cross Abstract: As atmospheric environmental prediction continues to improve, interpretable validation of pollution mechanisms and feedback processes has become a mai

agentsarxiv-cs-ai
9 Jun 2026
Research

TimpaTeks: Automatic In-place Text Sequence Modification via Diffusion Language Model Steering

DGX agent

arXiv:2606.08408v1 Announce Type: cross Abstract: We extend activation steering to diffusion language models (DLMs) and study a novel problem that arose due to the inference mechanism of DLMs: Modifyi

researcharxiv-cs-ai
9 Jun 2026
Research

TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech

DGX agent

arXiv:2606.09019v1 Announce Type: cross Abstract: Codec-based autoregressive (AR) speech language models have achieved strong text-to-speech (TTS) quality by modeling speech as sequences of discrete a

researcharxiv-cs-ai
9 Jun 2026
Agents

To Nuke or Not to Nuke: LLMs' (Missing) Ethical Reasoning and Actions in a High-Stakes Decision-Making Simulation

DGX agent

arXiv:2606.08310v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as long-horizon agents with decision-making capacities. While LLMs can show ethical competence on

agentsarxiv-cs-ai
9 Jun 2026
Research

Topological Neural Operators

DGX agent

arXiv:2606.09806v1 Announce Type: cross Abstract: We introduce Topological Neural Operators (TNOs), a principled framework for operator learning on cell complexes that lifts neural operators (NOs) fro

researcharxiv-cs-ai
9 Jun 2026
Safety

Toward autocorrection of chemical process flowsheets using large language models

DGX agent

arXiv:2312.02873v2 Announce Type: replace-cross Abstract: The process engineering domain widely uses Process Flow Diagrams (PFDs) and Process and Instrumentation Diagrams (P&IDs) to represent process

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Towards Long-Horizon Vessel Trajectory and Destination Forecasting with Reasoning Large Language Models

DGX agent

arXiv:2606.08633v1 Announce Type: new Abstract: Long-horizon maritime trajectory prediction is important for shipping management, logistics planning, and maritime risk analysis, yet month-level foreca

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…186187188189190…448
Next →