AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Tabular Foundation Models for Discrete Choice Estimation

DGX agent

arXiv:2607.13314v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) generate predictions on structured data via in-context learning, without task-specific estimation. We ask whether TFM

researcharxiv-cs-ai
16 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Cafe in Amsterdam: When the Incumbent Becomes the Oracle

DGX agent

arXiv:2607.13393v1 Announce Type: cross Abstract: A field can reformulate its computations freely exactly where its demand is stated independently of any incumbent implementation, and finds itself una

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

The Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Net Human-Agent Score (NHAS) in Autonomous Commerce

DGX agent

arXiv:2607.13998v1 Announce Type: cross Abstract: The rapid proliferation of Agentic Artificial Intelligence fundamentally disrupts traditional customer loyalty paradigms. As AI evolves from passive r

model-releasesarxiv-cs-ai
16 Jul 2026
Research

The Entanglement Wall: Activation-Space Probes as Risk Detectors, Not Context Adjudicators

DGX agent

arXiv:2607.13075v1 Announce Type: cross Abstract: Context can change whether a request is harmful without changing its topic or surface form. We ask whether residual-stream probes distinguish harmful

researcharxiv-cs-ai
16 Jul 2026
Tutorials

The Hitchhiker's Guide to Monoculture

DGX agent

arXiv:2607.13077v1 Announce Type: cross Abstract: Large language models (LLMs) often produce homogeneous outputs, raising concerns that AI coding assistants may lead to convergence in the software art

tutorialsarxiv-cs-ai
16 Jul 2026
Model Releases

The Perplexity Trap: When Patent Law Makes Human Writing Look Like AI

DGX agent

arXiv:2607.13044v1 Announce Type: cross Abstract: The European Patent Office (EPO) reported record filings in 2025, and the 2026 EPO Guidelines hold applicants strictly responsible for LLM-assisted co

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

The Refusal Residue: When Probes Catch Alignment Faking and When They Don't

DGX agent

arXiv:2607.13346v1 Announce Type: cross Abstract: Alignment faking is dangerous because a model can appear compliant under monitoring while preserving behavior it would reveal when unmonitored. When n

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

DGX agent

arXiv:2607.13612v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performa

safetyarxiv-cs-ai
16 Jul 2026
Research

TheBioCollection: Unified Pre-Training Scale LLM Corpus for Biology

DGX agent

arXiv:2607.08803v2 Announce Type: replace-cross Abstract: The push toward large language models for biology (BioLM) has created a need for training corpora that can endow models with a genuine underst

researcharxiv-cs-ai
16 Jul 2026
Research

Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases

DGX agent

arXiv:2607.13292v1 Announce Type: new Abstract: Autoformalization translates informal natural language into formal, machine-verifiable languages. While most work focuses on individual statements, real

researcharxiv-cs-ai
16 Jul 2026
Agents

Too Polite to Disagree: Understanding Sycophancy Propagation in Multi-Agent Systems

DGX agent

arXiv:2604.02668v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often exhibit sycophancy: agreement with user stance even when it conflicts with the model's opinion. While prior

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Traffic-Aware Randomized Smoothing for LLM-Based Network Intrusion Detection

DGX agent

arXiv:2607.13801v1 Announce Type: cross Abstract: Large language model (LLM)-based intrusion detection systems (IDS) are increasingly studied for security monitoring, yet their robustness against feas

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Transforming Rank: How Architecture Navigates the Spectral Pathologies of Depth

DGX agent

arXiv:2607.14018v1 Announce Type: cross Abstract: We investigate how each component of the Transformer feedforward block architecture design determines how much rank survives across depth at initializ

model-releasesarxiv-cs-ai
16 Jul 2026
Local Ai

TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling

DGX agent

arXiv:2607.13101v1 Announce Type: cross Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-ba

local-aiarxiv-cs-ai
16 Jul 2026
Model Releases

UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following

DGX agent

arXiv:2607.13621v1 Announce Type: new Abstract: Language-guided human following is an important capability for embodied agents, but existing benchmarks typically assume that the target person is visib

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streaming Systems

DGX agent

arXiv:2607.13048v1 Announce Type: cross Abstract: Streaming inference pipelines increasingly pair lightweight fast models with Large Language Models (LLMs) that provide rich semantic understanding at

safetyarxiv-cs-ai
16 Jul 2026
Research

Uniform Approximation of Functions with Asymmetric Growth and Decay by Deep Weighted Polynomials

DGX agent

arXiv:2506.21306v2 Announce Type: replace-cross Abstract: Functions that grow without bound on one side of the real line and decay to zero on the other cannot be approximated uniformly by ordinary pol

researcharxiv-cs-ai
16 Jul 2026
Agents

Unleashing Multimodal Large Language Models for Training-free HOI Detection in the Wild

DGX agent

arXiv:2607.13881v1 Announce Type: cross Abstract: Human-object interaction detection (HOID) has traditionally been formulated as a supervised detection problem over predefined interaction categories.

agentsarxiv-cs-ai
16 Jul 2026
Research

UTS at ELOQUENT 2026 Voight-Kampff: structural shifts in AI writing bypass state-of-the-art detectors

DGX agent

arXiv:2607.13565v1 Announce Type: cross Abstract: We investigate which language model evasion attacks survive state-of-the-art adversarial fine-tuning, developing strategies that sweep the top 5 posit

researcharxiv-cs-ai
16 Jul 2026
Applications

Verifier-Guided Twelve-Tone Composition: A Generate-Verify-Repair Harness for Symbolic Music Generation

DGX agent

arXiv:2607.11334v2 Announce Type: replace Abstract: Large language models can produce superficially legal twelve-tone scores that collapse into degenerate textures. We introduce a neuro-symbolic harne

applicationsarxiv-cs-ai
16 Jul 2026
Research

Verifying formulas for interventional distributions

DGX agent

arXiv:2607.13883v1 Announce Type: cross Abstract: We formalize verification in causal graphical models: deciding whether a given observational formula identifies a target interventional distribution.

researcharxiv-cs-ai
16 Jul 2026
Model Releases

WaterMoE: Expert-Routing-based Watermarking for High Fidelity and Efficiency

DGX agent

arXiv:2607.13099v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable success but raise growing concerns about content provenance and misuse, motivating the need for

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

What Models Express, Suppress, and Resist: Auditing Open-Weight LLMs with Persona Vectors

DGX agent

arXiv:2607.13162v1 Announce Type: cross Abstract: What a language model will and will not do is largely set during post-training, but which behaviors it expresses, hides, or resists is not revealed by

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

When Agents Disagree With Themselves: Behavioral Consistency as an Uncertainty Signal for LLM Agents

DGX agent

arXiv:2602.11619v2 Announce Type: replace Abstract: Running the same LLM agent on identical inputs yields 2.3-4.2 distinct action sequences per 10 runs; this behavioral variance constitutes a training

model-releasesarxiv-cs-ai
16 Jul 2026
Research

When Audio Separation Hurts Zero-Shot ASR: Evaluating SAM-Audio with Whisper on Bengali and English Speech

DGX agent

arXiv:2603.04710v2 Announce Type: replace-cross Abstract: Recent advances in automatic speech recognition (ASR) and speech enhancement have strengthened the common belief that cleaner audio should lea

researcharxiv-cs-ai
16 Jul 2026
Agents

When Bots Join the Team: Bot Adoption and the Institutional Fabric of Open-Source Software Projects

DGX agent

arXiv:2607.13679v1 Announce Type: new Abstract: AI agents are joining human teams, raising a basic question: when an automated agent becomes a regular participant, does group organization strengthen o

agentsarxiv-cs-ai
16 Jul 2026
Research

When is the combined load identifiable from a stress-intensity profile? A coupled forward-inverse study on SIFBench finite-element data

DGX agent

arXiv:2607.13074v1 Announce Type: cross Abstract: This work studies the inverse problem of recovering the relative magnitudes of the tension, bending, and bearing loads acting on a crack from its stre

researcharxiv-cs-ai
16 Jul 2026
Research

With Argus Eyes: Assessing Retrieval Gaps via Uncertainty Scoring to Detect and Remedy Retrieval Blind Spots

DGX agent

arXiv:2602.09616v2 Announce Type: replace-cross Abstract: Reliable retrieval-augmented generation (RAG) systems depend fundamentally on the retriever's ability to find relevant information. We show th

researcharxiv-cs-ai
16 Jul 2026
Model Releases

1D-Bench: A Benchmark for Iterative UI Code Generation with Visual Feedback in Real-World

DGX agent

arXiv:2602.18548v2 Announce Type: replace-cross Abstract: Design-to-code translates high-fidelity UI designs into executable front-end implementations, but progress remains hard to compare due to inco

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

A Comparative Analysis of Institutional and Course Generative AI Policies within Higher Education: Implications for Instruction in Computing Education

DGX agent

arXiv:2607.12296v1 Announce Type: cross Abstract: With the increased use of generative AI (GenAI) applications such as ChatGPT, higher education institutions (HEIs) have released a range of guidelines

model-releasesarxiv-cs-ai
15 Jul 2026
Local Ai

A Learning-Rate-Gated Failure of GRPO in a Small Language and Vision-Language Model Web Agent: A Controlled Null and Its Mechanism

DGX agent

arXiv:2607.12640v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards, and Group Relative Policy Optimization (GRPO) in particular, is now run routinely on a supervised checkp

local-aiarxiv-cs-ai
15 Jul 2026
Applications

A Longitudinal Analysis of Public Discourse on AI Ethics in Education Using Twitter Data

DGX agent

arXiv:2607.12295v1 Announce Type: cross Abstract: The rapid integration of artificial intelligence (AI) and generative AI (GenAI) into education presents significant opportunities to enhance teaching

applicationsarxiv-cs-ai
15 Jul 2026
Local Ai

A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study

DGX agent

arXiv:2607.12886v1 Announce Type: new Abstract: Clinical notes contain many of the signs and symptoms that bring patients to care, yet this information rarely reaches structured fields. Existing extra

local-aiarxiv-cs-ai
15 Jul 2026
Safety

A Neurosymbolic Approach to Natural Language Formalization and Verification

DGX agent

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits th

safetyarxiv-cs-ai
15 Jul 2026
Research

A Threshold Exceedance Framework for CBRN Uplift Evaluation in Frontier Language Models

DGX agent

arXiv:2607.12200v1 Announce Type: new Abstract: As frontier language models advance, policymakers and model developers need methods for assessing whether model access materially increases a non-expert

researcharxiv-cs-ai
15 Jul 2026
Safety

AAAI-26 Dual Submissions: Novel Challenges

DGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

ABot-N1: Toward a General Visual Language Navigation Foundation Model

DGX agent

arXiv:2607.10383v2 Announce Type: replace-cross Abstract: Visual Language Navigation foundation models aim to unify deep reasoning for grounded spatial decisions with broad versatility for diverse emb

model-releasesarxiv-cs-ai
15 Jul 2026
Research

Accelerating Masked Diffusion Large Language Models: A Survey of Efficient Inference Techniques

DGX agent

arXiv:2607.12829v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer a theoretical advantage in parallel generation over standard autoregressive models. However, parallel ge

researcharxiv-cs-ai
15 Jul 2026
Model Releases

Accepted Prefixes Are Not All You Need: A Negative Result on PEFT-Based Block-Diffusion Drafting

DGX agent

arXiv:2607.12422v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive language model inference by using a cheap drafter to propose multiple future tokens and a target model t

model-releasesarxiv-cs-ai
15 Jul 2026
Safety

Accuracy and Normalized Accuracy under Length Bias: Analysis, Guidelines, and a Bayesian Alternative

DGX agent

arXiv:2607.12767v1 Announce Type: new Abstract: Multiple-choice benchmarks that rank candidate completions by conditional log-probability suffer from a length bias: because log-probabilities sum over

safetyarxiv-cs-ai
15 Jul 2026
Tutorials

Action-Aware Generative Sequence Modeling for Short Video Recommendation

DGX agent

arXiv:2604.25834v2 Announce Type: replace Abstract: With the rapid development of the Internet, users have increasingly higher expectations for the recommendation accuracy of online content consumptio

tutorialsarxiv-cs-ai
15 Jul 2026
Research

Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter

DGX agent

arXiv:2607.10203v2 Announce Type: replace-cross Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better pre

researcharxiv-cs-ai
15 Jul 2026
Model Releases

Adaptive Testing for LLM Evaluation: A Psychometric Alternative to Static Benchmarks

DGX agent

arXiv:2511.04689v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) typically requires thousands of benchmark items, making the process expensive, slow, and increasingly

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Agent-Safety Evaluations as Load-Bearing Evidence: A Vendor-Neutral, Cross-Harness Reconstructability Metric

DGX agent

arXiv:2607.12469v1 Announce Type: cross Abstract: Many agent-safety evaluation results are not yet load-bearing evidence: identical nominal outcomes (task success, attack success, monitor scores) may

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

AgentCheck: A Reproduce-Intervene-Mitigate Workbench for LLM Agents over MCP

DGX agent

arXiv:2607.11098v2 Announce Type: replace-cross Abstract: Tool-using LLM agents are mostly evaluated assuming all tools work. When a tool times out, returns a week-stale value, or has its description

agentsarxiv-cs-ai
15 Jul 2026
Agents

Agentic Service-Oriented Computing: A Manifesto for the Next Frontier of Service-Oriented Computing

DGX agent

arXiv:2607.12619v1 Announce Type: new Abstract: The rapid emergence of LLM-powered autonomous and semi-autonomous agents is reshaping software systems from static, request-response components into goa

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

Agents Don't Just Agree, They Remember: Benchmarking Persistent Sycophancy in Stateful Personal Agents

DGX agent

arXiv:2607.10526v2 Announce Type: replace Abstract: Stateful personal agents increasingly maintain long-term user profiles, episodic memories, and reusable skills. This persistence turns conversationa

model-releasesarxiv-cs-ai
15 Jul 2026
Applications

An Empirical Analysis of Continual Learning for Heterogeneous Medical Visual Question Answering

DGX agent

arXiv:2607.12048v1 Announce Type: cross Abstract: Deploying medical visual question answering (MedVQA) systems in real-world clinical settings requires models that adapt to new clinical tasks without

applicationsarxiv-cs-ai
15 Jul 2026
← Previous
1…8687888990…448
Next →