AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

How Many Labels Are Enough? ALDA: Active Learning Deployment Advisor for Medical Image Classification

DGX agent

arXiv:2608.03511v1 Announce Type: cross Abstract: Active learning (AL) promises to reduce the cost of medical imaging projects by lowering the number of clinical labels required. However, practical de

researcharxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks

DGX agent

arXiv:2608.03502v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently shown strong capabilities in reasoning, planning, and tool-use, enabling new forms of autonomous agents. Howe

agentsarxiv-cs-ai
5 Aug 2026
Agents

HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM Agents

DGX agent

arXiv:2608.02650v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools to complete complex real-world tasks. However, reliable tool-use planning remains

agentsarxiv-cs-ai
5 Aug 2026
Research

Hypercubes, Hyperplanes, and Constraint-Induced Complexity Collapse in Atomic Concept Learning

DGX agent

arXiv:2608.02930v1 Announce Type: new Abstract: We revisit higher-arity atomic concept learning through the geometry of hypercubes and hyperplanes of ground instances. Our starting point is the observ

researcharxiv-cs-ai
5 Aug 2026
Model Releases

HyperFL: Query-Adaptive Representation Learning for Software Fault Localization

DGX agent

arXiv:2608.02967v1 Announce Type: cross Abstract: Software fault localization identifies the code locations responsible for reported issues and is a fundamental step toward automated debugging and pro

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Implementing Causal Perception: Competing SCMs and Situated Fairness

DGX agent

arXiv:2608.03917v1 Announce Type: new Abstract: Causal perception occurs when agents with competing Structural Causal Models (SCMs) of the same system infer different probability distributions, includ

safetyarxiv-cs-ai
5 Aug 2026
Safety

Improved Quantum Algorithms for Reinforcement Learning Under a Generative Model

DGX agent

arXiv:2608.02826v1 Announce Type: cross Abstract: Reinforcement learning is a subfield of machine learning that studies how an agent interacts with an environment in order to extract as large a reward

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

In-Context Collapse in Vision-Language Models and How to Mitigate it?

DGX agent

arXiv:2608.02830v1 Announce Type: cross Abstract: Many-shot in-context learning (ICL) lets vision-language models (VLMs) adapt from image--label demonstrations without weight updates, and is widely as

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

In-Context Pure Exploration in Continuous Decision Spaces

DGX agent

arXiv:2602.17976v2 Announce Type: replace-cross Abstract: In active sequential testing, also termed pure exploration, a learner is tasked with the goal to adaptively acquire information so as to ident

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Instruction Stacking Collapse: A Benchmark and the Capability-Dependent Value of Prompt Compilation

DGX agent

arXiv:2608.02639v1 Announce Type: cross Abstract: Production prompts rarely carry a single instruction. One system message may require valid JSON, a word limit, three citations, and a fixed tone at th

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent on a Public Blockchain

DGX agent

arXiv:2608.02986v1 Announce Type: cross Abstract: A software agent on a public blockchain accumulates authority and economic stakes, raising the engineering question of what makes it count as an indiv

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

DGX agent

arXiv:2608.03138v1 Announce Type: cross Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identif

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Interpretable Adaptive Sampling for LLM Test-Time Scaling

DGX agent

arXiv:2608.03961v1 Announce Type: new Abstract: Test-time scaling improves LLM reasoning by generating and aggregating multiple candidate answers, yet many pipelines use fixed per-query budgets that s

safetyarxiv-cs-ai
5 Aug 2026
Local Ai

Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes

DGX agent

arXiv:2608.02879v1 Announce Type: new Abstract: The widespread adoption of proprietary Large Language Models (LLMs) accessed strictly through closed APIs has created a critical challenge for responsib

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Intertemporal Preference Steering in Qwen3 via Contrastive Activation Addition

DGX agent

arXiv:2608.03892v1 Announce Type: new Abstract: We study linear representations of temporal horizon in the large language model Qwen3-32B and use them to change the model's time-related preferences, r

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning

DGX agent

arXiv:2507.14171v3 Announce Type: replace-cross Abstract: Importance-based structured pruning overwhelmingly relies on filter magnitude. This proxy is fundamentally flawed: due to scale invariance, fu

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autoformulation

DGX agent

arXiv:2608.02641v1 Announce Type: cross Abstract: Large language models (LLMs) can translate natural-language optimization problems into solver-ready formulations, but direct code generation is brittl

agentsarxiv-cs-ai
5 Aug 2026
Agents

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

DGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

agentsarxiv-cs-ai
5 Aug 2026
Applications

ISEE: Interactive Semantic Enrichment for Database Fields

DGX agent

arXiv:2608.02604v1 Announce Type: new Abstract: LLM-based agents are increasingly being deployed for data-related tasks, including data sense-making, exploration, and retrieval. However, their perform

applicationsarxiv-cs-ai
5 Aug 2026
Safety

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

DGX agent

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noi

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation

DGX agent

arXiv:2608.03782v1 Announce Type: new Abstract: Hallucination remains a critical challenge for developing trustworthy Multimodal Large Language Models (MLLMs). While existing benchmarks mainly focus o

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Knowing the Form, Not the Function: Automatically Auditing Answer--Authority Decoupling in Legal Benchmarks

DGX agent

arXiv:2608.02621v1 Announce Type: cross Abstract: Legal benchmarks typically score final answers even when models also state legal authority. We test whether answer correctness can serve as a proxy fo

model-releasesarxiv-cs-ai
5 Aug 2026
Applications

Large language models for partial differential equation workflows

DGX agent

arXiv:2608.03600v1 Announce Type: new Abstract: Partial differential equations (PDEs) become actionable in science and engineering not as isolated formulae, but as executable workflows that connect mo

applicationsarxiv-cs-ai
5 Aug 2026
Local Ai

Large Language Models provide support for the parallelogram theory of analogy

DGX agent

arXiv:2603.19066v2 Announce Type: replace-cross Abstract: Four-term word analogies (A:B::C:D) are classically modeled geometrically as parallelograms: adding the vector B-A+C produces D. Recent work s

local-aiarxiv-cs-ai
5 Aug 2026
Safety

LatentGuard: Efficient and Inspectable Latent Reasoning for LLM Safeguards

DGX agent

arXiv:2608.03838v1 Announce Type: new Abstract: Reasoning-based guard models improve LLM safeguards, but decoding explicit rationales for every interaction makes them costly to deploy. Although latent

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

LeanMem: Simple and Efficient Long-Term Memory for LLM Agents

DGX agent

arXiv:2608.03463v1 Announce Type: new Abstract: Long-term memory is essential for LLM-based agents to sustain interactions and reliably leverage distant history. However, existing memory systems typic

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Learning a Vector-Symbolic Model for Socio-Cultural Tasks

DGX agent

arXiv:2608.02807v1 Announce Type: cross Abstract: How can we better represent the impact of sociocultural structures on decision making in computational cognitive models? Modeling this impact requires

researcharxiv-cs-ai
5 Aug 2026
Safety

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents

DGX agent

arXiv:2608.03606v1 Announce Type: new Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence.

safetyarxiv-cs-ai
5 Aug 2026
Safety

Learning Molecular Representations from Cellular Phenotypes with Structure Preservation

DGX agent

arXiv:2608.02688v1 Announce Type: cross Abstract: Phenotypic drug discovery enables the discovery of functional relationships between molecular structures and cellular responses. However, existing mul

safetyarxiv-cs-ai
5 Aug 2026
Safety

Learning Music Style for Piano Arrangement Through Cross-Modal Bootstrapping

DGX agent

arXiv:2608.03050v1 Announce Type: cross Abstract: What is music style? Though often described using text labels such as 'swing,' 'classical,' or 'emotional,' the real style remains implicit and hidden

safetyarxiv-cs-ai
5 Aug 2026
Safety

Less Traffic, Better Outcomes: Competition-Aware Request Dispatch in Real-Time Ad Exchanges

DGX agent

arXiv:2608.03705v1 Announce Type: new Abstract: Real-time bidding (RTB) ad exchanges typically forward nearly all incoming requests to demand-side platforms (DSPs), even though only a small fraction r

safetyarxiv-cs-ai
5 Aug 2026
Applications

Leveraging System-Level Observations to Inform Bayesian Learning of Model Parameters for Quantitative Verification

DGX agent

arXiv:2608.03489v1 Announce Type: cross Abstract: Combining Bayesian learning and quantitative verification is a powerful toolset for analysing key quantitative properties of software systems, like re

applicationsarxiv-cs-ai
5 Aug 2026
Model Releases

Lightweight Chunk Selection for Mobile Retrieval-Augmented Generation

DGX agent

arXiv:2608.03148v1 Announce Type: cross Abstract: RAG improves the factual grounding of LLM by incorporating external knowledge, but deploying RAG on mobile and edge devices remains challenging becaus

model-releasesarxiv-cs-ai
5 Aug 2026
Hardware

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

DGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

hardwarearxiv-cs-ai
5 Aug 2026
Agents

LiveEvalBench: Toward Open-World Evaluation for Web Generation

DGX agent

arXiv:2608.03689v1 Announce Type: new Abstract: Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a sta

agentsarxiv-cs-ai
5 Aug 2026
Research

LLaDA MoE v2: Scaling Mixture-of-Experts Diffusion Language Models

DGX agent

arXiv:2608.03457v1 Announce Type: new Abstract: Diffusion language models (dLLMs) offer an alternative to autoregressive (AR) language modeling, yet the scaling behavior of Mixture-of-Experts (MoE) dL

researcharxiv-cs-ai
5 Aug 2026
Hardware

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

DGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

hardwarearxiv-cs-ai
5 Aug 2026
Model Releases

LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment

DGX agent

arXiv:2608.03020v1 Announce Type: new Abstract: Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Logic Before Language: Pre-pretraining on Formal Derivations Fosters Skill Acquisition and Compressibility

DGX agent

arXiv:2608.03930v1 Announce Type: cross Abstract: Pre-pretraining language models (LMs) on symbolic data can accelerate and improve natural language acquisition. However, existing pre-pretraining task

researcharxiv-cs-ai
5 Aug 2026
Applications

LogitScope: A Framework for Analyzing LLM Uncertainty Through Information Metrics

DGX agent

arXiv:2603.24929v2 Announce Type: replace Abstract: Understanding and quantifying uncertainty in large language model (LLM) outputs is critical for reliable deployment. However, traditional evaluation

applicationsarxiv-cs-ai
5 Aug 2026
Safety

Long-term Traffic Scene Prediction via Polynomial Representations in Autonomous Driving

DGX agent

arXiv:2608.03330v1 Announce Type: new Abstract: This thesis addresses fundamental challenges in traffic scene prediction for autonomous driving by introducing robust and computationally efficient mode

safetyarxiv-cs-ai
5 Aug 2026
Agents

MAFIA: Query-Only Memory Attacks via Probing and Factual Injection against Audited LLM Agents

DGX agent

arXiv:2608.03844v1 Announce Type: new Abstract: Memory-augmented LLM agents rely on rich context for long-horizon reasoning and acting, yet their memory modules expose a persistent attack surface for

agentsarxiv-cs-ai
5 Aug 2026
Safety

MambaTS: Improved Selective State Space Models for Long-term Time Series Forecasting

DGX agent

arXiv:2405.16440v2 Announce Type: replace-cross Abstract: In recent years, Transformers have become the de-facto architecture for long-term time series forecasting (LTSF), yet they face challenges ass

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

MDArena: Evaluating Coding Agents on Realistic Molecular Dynamics Workflows

DGX agent

arXiv:2608.02642v1 Announce Type: cross Abstract: Accelerating scientific discovery is among the most consequential applications of AI, and computational biomolecular simulation stands out as a partic

model-releasesarxiv-cs-ai
5 Aug 2026
Research

MDLMPE: Distribution Aware Positional Encoding for Masked Diffusion Language Models

DGX agent

arXiv:2608.03769v1 Announce Type: cross Abstract: Masked diffusion language models (MDLMs) enable parallel generation and bidirectional context modeling, but their positional context differs fundament

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation

DGX agent

arXiv:2608.00794v2 Announce Type: replace Abstract: Agentic AI evaluation pipelines produce benchmark scores that justify deployment decisions, safety certifications, and regulatory compliance claims.

model-releasesarxiv-cs-ai
5 Aug 2026
Tutorials

Measuring Explainer Stability via Attribution Separability

DGX agent

arXiv:2608.02697v1 Announce Type: cross Abstract: Attribution methods (AMs) assign an importance score to each feature and are widely adopted to explain black-box models. However, most methods can pro

tutorialsarxiv-cs-ai
5 Aug 2026
Research

Mechanism of Task-oriented Information Removal in In-context Learning

DGX agent

arXiv:2509.21012v4 Announce Type: replace-cross Abstract: In-context Learning (ICL) is an emerging few-shot learning paradigm based on modern Language Models (LMs), yet its inner mechanism remains unc

researcharxiv-cs-ai
5 Aug 2026
← Previous
1…4041424344…443
Next →