AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
Model Releases

The ATOM Report: Measuring the Open Language Model Ecosystem

DGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

model-releasesarxiv-cs-ai
10 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?

DGX agent

arXiv:2604.06436v2 Announce Type: cross Abstract: We prove that no continuous, utility-preserving wrapper defense-a function D: Xo X that preprocesses inputs before the model sees them-can make al

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning

DGX agent

arXiv:2604.06427v1 Announce Type: cross Abstract: The viability of chain-of-thought (CoT) monitoring hinges on models being unable to reason effectively in their latent representations. Yet little is

model-releasesarxiv-cs-ai
10 Apr 2026
Research

The Detection-Extraction Gap: Models Know the Answer Before They Can Say It

DGX agent

arXiv:2604.06613v2 Announce Type: cross Abstract: Modern reasoning models continue generating long after the answer is already determined. Across five model configurations, two families, and three ben

researcharxiv-cs-ai
10 Apr 2026
Safety

The End of the Foundation Model Era: Open-Weight Models, Sovereign AI, and Inference as Infrastructure

DGX agent

arXiv:2604.06217v1 Announce Type: cross Abstract: The foundation model era -- roughly 2020 to 2025 -- is over. The forces that defined it have inverted. Open source models have reached frontier perfor

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

The Geometry of Forgetting

DGX agent

arXiv:2604.06222v1 Announce Type: cross Abstract: Why do we forget? Why do we remember things that never happened? The conventional answer points to biological hardware. We propose a different one: ge

model-releasesarxiv-cs-ai
10 Apr 2026
Research

The Human Condition as Reflected in Contemporary Large Language Models

DGX agent

arXiv:2604.06206v1 Announce Type: cross Abstract: This study seeks to uncover evidence of a latent structure in evolved human culture as it is refracted through contemporary large language models (LLM

researcharxiv-cs-ai
10 Apr 2026
Model Releases

The Impact of Steering Large Language Models with Persona Vectors in Educational Applications

DGX agent

arXiv:2604.07102v1 Announce Type: cross Abstract: Activation-based steering can personalize large language models at inference time, but its effects in educational settings remain unclear. We study pe

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment

DGX agent

arXiv:2604.06377v1 Announce Type: cross Abstract: We investigate whether post-trained capabilities can be transferred across models without retraining, with a focus on transfer across different model

safetyarxiv-cs-ai
10 Apr 2026
Agents

The Planetary Cost of AI Acceleration, Part II: The 10th Planetary Boundary and the 6.5-Year Countdown

DGX agent

arXiv:2604.04956v2 Announce Type: replace-cross Abstract: The recent, super-exponential scaling of autonomous Large Language Model (LLM) agents signals a broader, fundamental paradigm shift from machi

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

The Stepwise Informativeness Assumption: Why are Entropy Dynamics and Reasoning Correlated in LLMs?

DGX agent

arXiv:2604.06192v1 Announce Type: cross Abstract: Recent work uses entropy-based signals at multiple representation levels to study reasoning in large language models, but the field remains largely em

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Traveling Thief Problem with Time Windows: Benchmarks and Heuristics

DGX agent

arXiv:2604.06724v1 Announce Type: cross Abstract: While traditional optimization problems were often studied in isolation, many real-world problems today require interdependence among multiple optimiz

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Thinking in Graphs with CoMAP: A Shared Visual Workspace for Designing Project-Based Learning

DGX agent

arXiv:2604.06200v1 Announce Type: cross Abstract: Designing project-based learning (PBL) demands managing highly interdependent components, a task that both traditional linear tools and purely convers

researcharxiv-cs-ai
10 Apr 2026
Safety

Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation

DGX agent

arXiv:2604.06205v1 Announce Type: cross Abstract: The growth of online platforms and user content requires strong content moderation systems that can handle complex inputs from various media types. Wh

safetyarxiv-cs-ai
10 Apr 2026
Research

Toward a Tractability Frontier for Exact Relevance Certification

DGX agent

arXiv:2604.07349v1 Announce Type: cross Abstract: Exact relevance certification asks which coordinates are necessary to determine the optimal action in a coordinate-structured decision problem. The tr

researcharxiv-cs-ai
10 Apr 2026
Model Releases

Toward a universal foundation model for graph-structured data

DGX agent

arXiv:2604.06391v1 Announce Type: cross Abstract: Graphs are a central representation in biomedical research, capturing molecular interaction networks, gene regulatory circuits, cell--cell communicati

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

DGX agent

arXiv:2505.22976v2 Announce Type: replace-cross Abstract: The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Toward Reducing Unproductive Container Moves: Predicting Service Requirements and Dwell Times

DGX agent

arXiv:2604.06251v1 Announce Type: new Abstract: This article presents the results of a data science study conducted at a container terminal, aimed at reducing unproductive container moves through the

researcharxiv-cs-ai
10 Apr 2026
Safety

Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation

DGX agent

arXiv:2604.06831v1 Announce Type: cross Abstract: Current LLM-based services typically require users to submit raw text regardless of its sensitivity. While intuitive, such practice introduces substan

safetyarxiv-cs-ai
10 Apr 2026
Safety

Towards provable probabilistic safety for scalable embodied AI systems

DGX agent

arXiv:2506.05171v3 Announce Type: replace-cross Abstract: Embodied AI systems, comprising AI models and physical plants, are increasingly prevalent across various applications. Due to the rarity of sy

safetyarxiv-cs-ai
10 Apr 2026
Agents

Towards Resilient Intrusion Detection in CubeSats: Challenges, TinyML Solutions, and Future Directions

DGX agent

arXiv:2604.06411v1 Announce Type: cross Abstract: CubeSats have revolutionized access to space by providing affordable and accessible platforms for research and education. However, their reliance on C

agentsarxiv-cs-ai
10 Apr 2026
Safety

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

DGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

DGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

DGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

model-releasesarxiv-cs-ai
10 Apr 2026
Applications

TREASURE: The Visa Payment Foundation Model for High-Volume Transaction Understanding

DGX agent

arXiv:2511.19693v3 Announce Type: replace-cross Abstract: Payment networks form the backbone of modern commerce, generating high volumes of transaction records from daily activities. Properly modeling

applicationsarxiv-cs-ai
10 Apr 2026
Agents

TurboAgent: An LLM-Driven Autonomous Multi-Agent Framework for Turbomachinery Aerodynamic Design

DGX agent

arXiv:2604.06747v2 Announce Type: new Abstract: The aerodynamic design of turbomachinery is a complex and tightly coupled multi-stage process involving geometry generation, performance prediction, opt

agentsarxiv-cs-ai
10 Apr 2026
Safety

TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.06610v1 Announce Type: cross Abstract: Decentralised online learning enables runtime adaptation in cyber-physical multi-agent systems, but when operating conditions change, learned policies

safetyarxiv-cs-ai
10 Apr 2026
Agents

UI-AGILE: Advancing GUI Agents with Effective Reinforcement Learning and Precise Inference-Time Grounding

DGX agent

arXiv:2507.22025v4 Announce Type: replace Abstract: The emergence of Multimodal Large Language Models (MLLMs) has driven significant advances in Graphical User Interface (GUI) agent capabilities. Neve

agentsarxiv-cs-ai
10 Apr 2026
Agents

Uncertainty Estimation for Deep Reconstruction in Actuatic Disaster Scenarios with Autonomous Vehicles

DGX agent

arXiv:2604.06387v1 Announce Type: cross Abstract: Accurate reconstruction of environmental scalar fields from sparse onboard observations is essential for autonomous vehicles engaged in aquatic monito

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

Unifying Speech Editing Detection and Content Localization via Prior-Enhanced Audio LLMs

DGX agent

arXiv:2601.21463v2 Announce Type: replace-cross Abstract: Existing speech editing detection (SED) datasets are predominantly constructed using manual splicing or limited editing operations, resulting

model-releasesarxiv-cs-ai
10 Apr 2026
Applications

Unsupervised Neural Network for Automated Classification of Surgical Urgency Levels in Medical Transcriptions

DGX agent

arXiv:2604.06214v1 Announce Type: cross Abstract: Efficient classification of surgical procedures by urgency is paramount to optimize patient care and resource allocation within healthcare systems. Th

applicationsarxiv-cs-ai
10 Apr 2026
Safety

URMF: Uncertainty-aware Robust Multimodal Fusion for Multimodal Sarcasm Detection

DGX agent

arXiv:2604.06728v1 Announce Type: cross Abstract: Multimodal sarcasm detection (MSD) aims to identify sarcastic intent from semantic incongruity between text and image. Although recent methods have im

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Validated Intent Compilation for Constrained Routing in LEO Mega-Constellations

DGX agent

arXiv:2604.07264v1 Announce Type: cross Abstract: Operating LEO mega-constellations requires translating high-level operator intents ('reroute financial traffic away from polar links under 80 ms') int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics

DGX agent

arXiv:2604.06182v1 Announce Type: cross Abstract: Existing online benchmarks for mobile GUI agents remain largely app-centric and task-homogeneous, failing to reflect the diversity and instability of

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

VisCoder2: Building Multi-Language Visualization Coding Agents

DGX agent

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

VisionClaw: Always-On AI Agents through Smart Glasses

DGX agent

arXiv:2604.03486v2 Announce Type: replace-cross Abstract: We present VisionClaw, an always-on wearable AI agent that integrates live egocentric perception with agentic task execution. Running on Meta

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

Weakly Supervised Distillation of Hallucination Signals into Transformer Representations

DGX agent

arXiv:2604.06277v1 Announce Type: new Abstract: Existing hallucination detection methods for large language models (LLMs) rely on external verification at inference time, requiring gold answers, retri

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search

DGX agent

arXiv:2604.06177v1 Announce Type: cross Abstract: Specialized web tasks in finance, biomedicine, and pharmaceuticals remain challenging due to missing domain priors: queries drift, evidence is noisy,

safetyarxiv-cs-ai
10 Apr 2026
Safety

WebSP-Eval: Evaluating Web Agents on Website Security and Privacy Tasks

DGX agent

arXiv:2604.06367v1 Announce Type: cross Abstract: Web agents automate browser tasks, ranging from simple form completion to complex workflows like ordering groceries. While current benchmarks evaluate

safetyarxiv-cs-ai
10 Apr 2026
Safety

What Makes an Ideal Quote? Recommending 'Unexpected yet Rational' Quotations via Novelty

DGX agent

arXiv:2602.22220v2 Announce Type: replace-cross Abstract: Quotation recommendation aims to enrich writing by suggesting quotes that complement a given context, yet existing systems mostly optimize sur

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

What's Missing in Screen-to-Action? Towards a UI-in-the-Loop Paradigm for Multimodal GUI Reasoning

DGX agent

arXiv:2604.06995v1 Announce Type: new Abstract: Existing Graphical User Interface (GUI) reasoning tasks remain challenging, particularly in UI understanding. Current methods typically rely on direct s

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't

DGX agent

arXiv:2604.06422v1 Announce Type: cross Abstract: Understanding when Vision-Language Models (VLMs) will behave unexpectedly, whether models can reliably predict their own behavior, and if models adher

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

Working Paper: Towards a Category-theoretic Comparative Framework for Artificial General Intelligence

DGX agent

arXiv:2603.28906v2 Announce Type: replace Abstract: AGI has become the Holly Grail of AI with the promise of level intelligence and the major Tech companies around the world are investing unprecedente

agentsarxiv-cs-ai
10 Apr 2026
Research

WRAP++: Web discoveRy Amplified Pretraining

DGX agent

arXiv:2604.06829v2 Announce Type: cross Abstract: Synthetic data rephrasing has emerged as a powerful technique for enhancing knowledge acquisition during large language model (LLM) pretraining. Howev

researcharxiv-cs-ai
10 Apr 2026
Research

XR-CareerAssist: An Immersive Platform for Personalised Career Guidance Leveraging Extended Reality and Multimodal AI

DGX agent

arXiv:2604.06901v1 Announce Type: cross Abstract: Conventional career guidance platforms rely on static, text-driven interfaces that struggle to engage users or deliver personalised, evidence-based in

researcharxiv-cs-ai
10 Apr 2026
Research

Zatom-1: A Multimodal Flow Foundation Model for 3D Molecules and Materials

DGX agent

arXiv:2602.22251v3 Announce Type: replace-cross Abstract: General-purpose 3D chemical modeling encompasses molecules and materials, requiring both generative and predictive capabilities. However, most

researcharxiv-cs-ai
10 Apr 2026
Syntheses

Wiki Lint Report — 2026-04-26

DGX agent

Automated lint: 44 errors, 10 warnings, 3 info

linthealth-checkautomated
26 Apr 2026
Syntheses

Wiki Lint Report — 2026-04-19

DGX agent

Automated lint: 43 errors, 9 warnings, 3 info

linthealth-checkautomated
19 Apr 2026
← Previous
1…435436437438
Next →