AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
Model Releases

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

DGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Robust Renal Mass Segmentation on CT: A Validation Study of an AI-Based Framework

DGX agent

arXiv:2505.07573v2 Announce Type: replace-cross Abstract: Renal mass segmentation has important potential to enhance the clinical workflow, especially in settings requiring quantitative assessments. K

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?

DGX agent

arXiv:2606.08063v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in visual understanding, yet their performance degrades significantly un

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

DGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

DGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

Rule-based autocorrection of Piping and Instrumentation Diagrams (P&IDs) on graphs

DGX agent

arXiv:2502.18493v2 Announce Type: replace-cross Abstract: A piping and instrumentation diagram (P&ID) is a central reference document in chemical process engineering. Currently, chemical engineers man

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

RunAgent SuperBrowser: A Theory of Autonomous Web Navigation Grounded in Human Browsing Behaviour

DGX agent

arXiv:2606.09399v1 Announce Type: new Abstract: We present SUPERBROWSER, an autonomous web-navigation agent designed against a single guiding hypothesis: a web agent should browse the way a person bro

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Safe-RULE: Safe Reinforcement UnLEarning

DGX agent

arXiv:2606.09559v1 Announce Type: cross Abstract: Offline safe reinforcement learning (Safe RL) enables policy learning without online interactions, making it suitable for safety-critical systems such

model-releasesarxiv-cs-ai
9 Jun 2026
Research

SafeECGMatch: Calibration-Aware Joint Frequency and Time Space Semi-Supervised Learning for Open-Set ECG Classification

DGX agent

arXiv:2606.08037v1 Announce Type: cross Abstract: Electrocardiogram (ECG) classification models often suffer from severe label scarcity, making semi-supervised learning (SSL) an attractive strategy fo

researcharxiv-cs-ai
9 Jun 2026
Model Releases

SafeRun: Enabling Determinism in LLM Planning for Running

DGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators

DGX agent

arXiv:2606.07874v1 Announce Type: new Abstract: LLMs-as-judges are the only way to evaluate safety at scale. Despite their importance, LLM-judges themselves are rarely evaluated beyond human agreement

safetyarxiv-cs-ai
9 Jun 2026
Agents

SAGE: An LLM-driven Self Reflective Agentic Framework for Fraud Detection

DGX agent

arXiv:2606.08146v1 Announce Type: new Abstract: Fraud detection in payment, e-commerce, and telecommunications systems requires accuracy at the individual level, robustness under severe class imbalanc

agentsarxiv-cs-ai
9 Jun 2026
Research

SAGE: Shape-Adapting Gated Experts for Adaptive Histopathology Image Segmentation

DGX agent

arXiv:2511.18493v4 Announce Type: replace-cross Abstract: The significant variability in cell size and shape continues to pose a major obstacle in computer-assisted cancer detection on gigapixel Whole

researcharxiv-cs-ai
9 Jun 2026
Local Ai

SAILS: Surrogate-based Analysis of Interactions via Local Effect Smooths

DGX agent

arXiv:2606.09404v1 Announce Type: cross Abstract: Feature interactions drive much of the predictive power of machine learning models, yet existing explanation methods only detect and quantify interact

local-aiarxiv-cs-ai
9 Jun 2026
Tutorials

Sample-Efficient LLM-Based Detection of Malicious Web Server Logs with Forensically Explainable Reasoning

DGX agent

arXiv:2606.08649v1 Announce Type: cross Abstract: Forensic analysis of web server logs demands both accurate detection and human-readable explanations that can satisfy legal requirements. We present C

tutorialsarxiv-cs-ai
9 Jun 2026
Safety

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

DGX agent

arXiv:2606.07602v1 Announce Type: cross Abstract: LLM-based LEGO assembly generation requires both semantic grounding and physical feasibility. We identify a data-induced failure mode, PhysHack, in wh

safetyarxiv-cs-ai
9 Jun 2026
Safety

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

DGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Scaffold Effects on GAIA: A Controlled Comparison

DGX agent

arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ScaleSweep: Accurate NVFP4 Post-Training Quantization of LLMs via Block Scale Initialization

DGX agent

arXiv:2606.07618v1 Announce Type: cross Abstract: NVFP4 is a recently introduced hardware-supported FP4 format that improves the fidelity of 4-bit quantization through fine-grained block scales. Howev

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Scaling Decision-Focused Learning to Large Problems with Lagrangian Decomposition

DGX agent

arXiv:2606.08797v1 Announce Type: cross Abstract: Decision-focused learning has shown great promise for addressing predict-then-optimize problems, particularly in the presence of under-specified model

researcharxiv-cs-ai
9 Jun 2026
Safety

Scaling Neural Network Verification with Tensor Parallelism and Fully Sharded Data Parallelism

DGX agent

arXiv:2606.09377v1 Announce Type: cross Abstract: Formal neural network verification -- proving that a network satisfies safety properties for all inputs in a specified domain -- is bounded in practic

safetyarxiv-cs-ai
9 Jun 2026
Research

Scaling Participation in Modular AI Systems

DGX agent

arXiv:2606.07812v1 Announce Type: new Abstract: Humanity is a mosaic of multifaceted talents and needs, and any truly intelligent AI must reflect that richness. Yet the LLMs used by all are built by t

researcharxiv-cs-ai
9 Jun 2026
Model Releases

SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration

DGX agent

arXiv:2606.08402v1 Announce Type: cross Abstract: Generating complete 3D scenes from a single image requires inferring globally consistent geometry, object relationships, and environmental context fro

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

DGX agent

arXiv:2606.08034v1 Announce Type: cross Abstract: Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing s

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

SciTrace: Trajectory-Aware Safety Reasoning for Scientific Discovery Agents

DGX agent

arXiv:2606.08234v1 Announce Type: new Abstract: LLM-based scientific agents have shown strong capacity for autonomous research, yet their safety layers remain structurally divorced from core reasoning

safetyarxiv-cs-ai
9 Jun 2026
Agents

SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Research

DGX agent

arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose context demands can grow without bound, yet model

agentsarxiv-cs-ai
9 Jun 2026
Safety

SecureClaw: Clawing Back Control of LLM Agents

DGX agent

arXiv:2606.09549v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents face two distinct security failures: unauthorized external actions and exposure of sensitive plaintext in

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

DGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Seeing is Believing: Aligning Prompt Rewriting with Visual Anchors for Text-to-Image Generation

DGX agent

arXiv:2606.08492v1 Announce Type: cross Abstract: Despite the impressive capabilities of text-to-image (T2I) models, an intent-generation gap often persists due to the brevity and ambiguity of user pr

researcharxiv-cs-ai
9 Jun 2026
Local Ai

Seeing the Hivemind: A Consensus-Aware Interaction Technique for Mitigating AI Homogenization

DGX agent

arXiv:2606.09587v1 Announce Type: cross Abstract: People are increasingly using AI for creative tasks such as writing. While adoption continues to grow, this form of use risks undermining individual c

local-aiarxiv-cs-ai
9 Jun 2026
Safety

SEF-CLGC at SemEval-2026 Task 11: Logical Notation Impact on Language Model Performance

DGX agent

arXiv:2606.09157v1 Announce Type: cross Abstract: This paper revisits our pipeline called Syllogistic Evaluation Framework-Common Logic Grammar Construction (SEF-CLGC). We combine formal logical notat

safetyarxiv-cs-ai
9 Jun 2026
Research

Segment-level Tree Search for Long Meeting Document Summarization

DGX agent

arXiv:2606.08445v1 Announce Type: cross Abstract: Meeting documents are challenging to summarize due to their length and complex conversational structure. Existing approaches typically adopt multi-sta

researcharxiv-cs-ai
9 Jun 2026
Applications

Selecting New Measurement Locations to Diversify Traffic-Pattern Coverage: A Real-World Evaluation for Total Traffic Volume Estimation

DGX agent

arXiv:2606.07556v1 Announce Type: cross Abstract: Accurate measurement of traffic volumes and flows is vital for modern intelligent transportation. However, despite recent technological advances in se

applicationsarxiv-cs-ai
9 Jun 2026
Safety

Self-Evolving Scientific Agent Discovers Generalizable Physically-Reasoned Fluid Control

DGX agent

arXiv:2606.08405v1 Announce Type: new Abstract: While data-intensive deep reinforcement learning can optimize complex control policies, scientific discovery in physical systems fundamentally requires

safetyarxiv-cs-ai
9 Jun 2026
Research

Self-Explainability in Self-Adaptive and Self-Organising Systems: Status and Research Directions

DGX agent

arXiv:2606.09568v1 Announce Type: new Abstract: The growing complexity of self-adaptive and self-organising systems, fuelled by advances in Artificial Intelligence (AI), has made them increasingly dif

researcharxiv-cs-ai
9 Jun 2026
Agents

Self-Paced Curriculum Reinforcement Learning for Autonomous Superbike Racing in Simulation

DGX agent

arXiv:2606.09236v1 Announce Type: cross Abstract: Autonomous Racing has seen remarkable progress through deep Reinforcement Learning (RL), primarily for four-wheeled vehicles. However, motorbikes intr

agentsarxiv-cs-ai
9 Jun 2026
Research

Self-Supervised Vision Transformers for CBCT-Based Detection of Temporomandibular Joint Osteoarthritis

DGX agent

arXiv:2606.08364v1 Announce Type: cross Abstract: Temporomandibular joint osteoarthritis (TMJ OA) is a prevalent degenerative condition whose osseous changes are often subtle on cone-beam CT (CBCT), m

researcharxiv-cs-ai
9 Jun 2026
Research

Semantic Cache Distillation: Efficient State Transfer via Reuse and Selective Patching

DGX agent

arXiv:2606.07684v1 Announce Type: cross Abstract: Disaggregated serving alleviates memory bottlenecks in Large Language Model (LLM) inference but creates a severe communication bottleneck: transmittin

researcharxiv-cs-ai
9 Jun 2026
Safety

Semantic Quorum Assurance: Collective Certification for Non-Deterministic AI Infrastructure

DGX agent

arXiv:2606.08021v1 Announce Type: cross Abstract: As large language model (LLM) agents are integrated into autonomous cloud operations, distributed systems face a semantic reliability problem: propose

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

SENTRY: Statistical Reliability Analysis of Vision Transformers Under Soft Errors

DGX agent

arXiv:2606.07620v1 Announce Type: cross Abstract: With the growth of Vision Transformers in safety-critical domains like autonomous systems and medical imaging, ensuring their reliability against soft

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Seq103: A Unified Neuroevolution Framework for Compact Sequence Architecture Discovery

DGX agent

arXiv:2606.07664v1 Announce Type: cross Abstract: Neuroevolution is a representative neural architecture search paradigm that evolves both network topology and weights through evolutionary algorithms.

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Set-Based Transformer for Atmospheric Compensation in Standoff LWIR Hyperspectral Imaging

DGX agent

arXiv:2606.08324v1 Announce Type: cross Abstract: Passive long-wave infrared (LWIR) hyperspectral imaging under a standoff geometry depends on atmospheric absorption and emission, as well as reflected

researcharxiv-cs-ai
9 Jun 2026
Safety

sGPO: Trading Inference FLOPs for Training Efficiency in RLVR

DGX agent

arXiv:2606.08854v1 Announce Type: cross Abstract: Standard Reinforcement Learning with Verifiable Rewards (RLVR) training allocates a fixed rollout budget to every query, without regard for what each

safetyarxiv-cs-ai
9 Jun 2026
Agents

Shape Formation for the Cooperative Transportation of Arbitrary Objects Using Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.09610v1 Announce Type: cross Abstract: Cooperative object transportation is essential in numerous domains, including industrial to domestic services. A popular transportation strategy is to

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Shared Latent Structures Enable Unified Backdoor Detection and Mitigation in LLMs

DGX agent

arXiv:2606.07963v1 Announce Type: new Abstract: Backdoor attacks in large language models (LLMs) are often treated as isolated trigger-response failures, motivating defenses tailored to specific trigg

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…189190191192193…452
Next →