AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

DGX agent

arXiv:2607.22529v1 Announce Type: new Abstract: LLM training is shifting from manual design and annotation to interaction-driven self-evolution. However, existing self-evolutionary methods face a fund

model-releasesarxiv-cs-cl
27 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Unboxing Diffusion Models for the Arts: Interactive Model Bending and Practice-Based Explainability

DGX agent

arXiv:2607.22428v1 Announce Type: cross Abstract: Explainable AI (XAI) in creative practice can be less about technocentric explanation and more about enabling artists to inspect modify and debug mode

researcharxiv-cs-lg
27 Jul 2026
Research

A Situational Speech Synthesizer for Yoruba: System Design, Phonological Rule Architecture, and Orthographic Extensions for Contour

DGX agent

arXiv:2607.18317v2 Announce Type: replace-cross Abstract: We present TTSYoruba, a rule-based concatenative diphone speech synthesizer for Yoruba, deployed at online as part of the YorubaName.com open

researcharxiv-cs-cl
24 Jul 2026
Research

Chess_db: A framework for working with large chess game datasets

DGX agent

arXiv:2607.21195v1 Announce Type: cross Abstract: Chess is a two player strategic game that is embedded in classical AI culture as it was once the frontier for intelligent behaviour. There was the sil

researcharxiv-cs-ai
24 Jul 2026
Research

ClickGuard: Detecting and Spoiling Clickbait News with Informativeness Measures and Large Language Models

DGX agent

arXiv:2607.20463v1 Announce Type: new Abstract: This paper presents an AI-driven browser extension that identifies clickbait to help users avoid misleading Internet articles. Moving beyond traditional

researcharxiv-cs-ai
24 Jul 2026
Research

Distinguishing Artificial from Authentic: Evaluating LLMs for Detecting LLM-Generated Content

DGX agent

arXiv:2607.20446v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly used by students to generate natural language responses and program code, there is growing interest in

researcharxiv-cs-cl
24 Jul 2026
Model Releases

Domyn-Small: A European 10B Reasoning Language Model

DGX agent

arXiv:2607.20448v1 Announce Type: new Abstract: We introduce Domyn-Small, a 10-billion-parameter open-weight reasoning language model released under the MIT license. Domyn-Small is the product of an i

model-releasesarxiv-cs-cl
24 Jul 2026
Tutorials

EZASP - Facilitating the Usage of ASP

DGX agent

arXiv:2603.26863v2 Announce Type: replace-cross Abstract: Answer Set Programming (ASP) is a declarative programming language used for modeling and solving complex combinatorial problems. It has been s

tutorialsarxiv-cs-ai
24 Jul 2026
Research

Foundation-model-guided radiogenomic discovery linking cancer genomes to cancer scans

DGX agent

arXiv:2607.20583v1 Announce Type: cross Abstract: The function of many genes is still unknown, and conventional driver-discovery methods, which rely on how frequently a gene is mutated, cannot assess

researcharxiv-cs-ai
24 Jul 2026
Safety

Human-Inspired Framework for Robotic Craniotomy: Integrating Multimodal Fusion and Adaptive Trajectory Adjustment

DGX agent

arXiv:2607.21058v1 Announce Type: new Abstract: Manual craniotomy is a high-risk, skill-dependent procedure associated with surgeon fatigue and potential dural injury. While robotic approaches have im

safetyarxiv-cs-ro
24 Jul 2026
Model Releases

ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders

DGX agent

arXiv:2607.21217v1 Announce Type: new Abstract: The recent emergence of vibe-coding workflows is changing what coding agents are expected to do. Instead of merely completing code under fully specified

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit

DGX agent

arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases

researcharxiv-cs-ai
24 Jul 2026
Applications

Learning to Detect UI Principle Violations via Reinforcement Learning

DGX agent

arXiv:2607.20690v1 Announce Type: new Abstract: Small language models and coding agents increasingly generate web front-end code, yet their outputs are typically evaluated primarily for functional cor

applicationsarxiv-cs-cl
24 Jul 2026
Model Releases

LegalCiteTrust: Benchmarking Citation Trustworthiness in Chinese Long-Form Legal Research Reports

DGX agent

arXiv:2607.20872v1 Announce Type: new Abstract: Long-form legal research reports increasingly rely on LLMs and agentic research systems, but their reliability depends not only on answering the task, b

model-releasesarxiv-cs-cl
24 Jul 2026
Safety

Multivariate Planar Curves: A Statistical Framework for Shape Analysis in Images

DGX agent

arXiv:2508.11780v3 Announce Type: replace-cross Abstract: Recent developments in computer vision have made segmented images widely available across many domains, such as medicine, where segmented radi

safetyarxiv-cs-cv
24 Jul 2026
Model Releases

Position Bias is Hidden Behind Ceiling Effects: A Permutation Diagnostic for LLM Benchmarks

DGX agent

arXiv:2607.20864v1 Announce Type: cross Abstract: Position bias in multiple-choice LLM evaluation is widely cited as a confound in capability comparisons, but published measurements rely on single ans

model-releasesarxiv-cs-cl
24 Jul 2026
Applications

Position: Natural Language Should Not Fully Replace Formal Languages

DGX agent

arXiv:2607.20432v1 Announce Type: new Abstract: Recent advances in large language models and their widespread adoption have prompted claims that natural language could entirely replace formal language

applicationsarxiv-cs-cl
24 Jul 2026
Research

Post-Hoc Reasoning in Chain of Thought: Decoding and Steering Pre-Committed Answers

DGX agent

arXiv:2603.01437v2 Announce Type: replace Abstract: As chain of thought (CoT) has become central to scaling reasoning capabilities in large language models (LLMs), it has also emerged as a promising t

researcharxiv-cs-ai
24 Jul 2026
Research

Progressive Cramming: Reliable Token Compression and What It Reveals

DGX agent

arXiv:2607.21231v1 Announce Type: new Abstract: Token cramming compresses sequences into learned embeddings with near-perfect reconstruction, but fixed token budgets and 99% accuracy thresholds leave

researcharxiv-cs-cl
24 Jul 2026
Safety

Regulating autonomous and agentic AI

DGX agent

arXiv:2607.21345v1 Announce Type: new Abstract: Regulating activities where regulatees use autonomous and agentic AI is challenging. Regulatory assumptions about regulatee knowledge and control no lon

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

RUMBA: Russian User Memory Benchmark

DGX agent

arXiv:2607.21447v1 Announce Type: cross Abstract: The ability to handle long-term memory in LLMs is becoming increasingly critical, yet existing benchmarks remain English-centric and rely on aggregate

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The Hidden Footprint: Making Storage a First-Class Metric for LLM Agent Evaluation

DGX agent

arXiv:2607.11149v3 Announce Type: replace Abstract: LLM agent benchmarks measure task completion, reliability, and inference cost, but not the persistent data an agent run leaves on disk, including lo

model-releasesarxiv-cs-ai
24 Jul 2026
Local Ai

Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry

DGX agent

arXiv:2607.21495v1 Announce Type: new Abstract: AI agents are increasingly created inside organizations by non-engineering users through low-code, no-code, and conversational development environments.

local-aiarxiv-cs-ai
24 Jul 2026
Research

Transition-Related Potentials as Markers of Narrative Comprehension in Continuous EEG

DGX agent

arXiv:2607.20720v1 Announce Type: cross Abstract: Harnessing the potential of electroencephalography (EEG) for brain research is fundamentally limited by intrinsic noise and the diffuse projection of

researcharxiv-cs-ai
24 Jul 2026
Research

Unsupervised Metal Artifact Reduction in Dental CBCT using Fine-tuned Cycle-Consistent Adversarial Networks

DGX agent

arXiv:2607.20977v1 Announce Type: new Abstract: Metal artifacts generated by dental implants significantly degrade cone-beam computed tomography (CBCT) volumes, obscuring critical anatomical structure

researcharxiv-cs-cv
24 Jul 2026
Model Releases

When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion

DGX agent

arXiv:2607.20543v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve one-sample accuracy while making a model worse under repeated sampling. We study thi

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Writhe-Based Polymer Link Classification Using Machine Learning

DGX agent

arXiv:2607.20657v1 Announce Type: cross Abstract: Unique and rapid classification of knots and links is an open mathematical problem that is relevant to a range of (bio)physical systems, including pol

researcharxiv-cs-lg
24 Jul 2026
Research

A Bayesian Framework for Built-in Input Dimension Reduction for Gaussian Process Modeling

DGX agent

arXiv:2607.19498v1 Announce Type: cross Abstract: Gaussian process (GP) modeling is widely used in computational science and engineering. However, fitting a GP to high-dimensional inputs remains chall

researcharxiv-cs-lg
23 Jul 2026
Agents

Agent-Centric Animal Pose Forecasting

DGX agent

arXiv:2607.19548v1 Announce Type: new Abstract: Understanding animal behavior at an algorithmic level -- what animals attend to, how they form internal models and plans, and how this maps to action --

agentsarxiv-cs-lg
23 Jul 2026
Safety

Agentic Real2Sim: Physics-based World Modeling with Vision-Language Agents

DGX agent

arXiv:2607.19190v2 Announce Type: replace-cross Abstract: Real-to-sim conversion for robotic interaction with objects remains labor-intensive because it requires more than visual reconstruction: a str

safetyarxiv-cs-ai
23 Jul 2026
Research

Asymptotically Optimal Regret for Reinforcement Learning without Horizon Dependence

DGX agent

arXiv:2607.19854v1 Announce Type: new Abstract: We study horizon-free regret minimization for finite-horizon time-homogeneous tabular Markov decision processes with S states, A actions, horizon H, and

researcharxiv-cs-lg
23 Jul 2026
Agents

AutoVSR: Automatic Visual-to-Symbolic Reasoning for Symbolic Expression Generation from Circuit Schematic

DGX agent

arXiv:2607.11338v2 Announce Type: replace Abstract: Symbolic expressions can effectively characterize and predict circuit behavior, but deriving them directly from circuit schematics is challenging. T

agentsarxiv-cs-ai
23 Jul 2026
Research

Building Fast, Evaluating Slow: Pipeline Choices Dominate Autointerpretability Score Variance

DGX agent

arXiv:2607.19386v1 Announce Type: cross Abstract: Cross-paper comparison of sparse autoencoder (SAE) interpretability often relies on autointerpretability scores. In this evaluation pipeline, a langua

researcharxiv-cs-cl
23 Jul 2026
Applications

Challenges of Explainability in Continual Learning for Time Series Forecasting

DGX agent

arXiv:2607.19382v1 Announce Type: cross Abstract: Deep learning models have shown strong potential for time series forecasting, yet their deployment in real-world environmental monitoring remains chal

applicationsarxiv-cs-ai
23 Jul 2026
Model Releases

Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail Humanoids

DGX agent

arXiv:2607.20345v1 Announce Type: cross Abstract: Closing the gap between benchmark performance and reliable real-world operation remains a central challenge for Vision-Language-Action (VLA) humanoid

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation

DGX agent

arXiv:2607.15434v3 Announce Type: replace-cross Abstract: Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

DGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

D2VBench: Benchmarking Large Language Models with Value Dilemmas in Daily Scenarios

DGX agent

arXiv:2607.19834v1 Announce Type: new Abstract: With the wide application of large language models (LLMs) in real-world scenarios, the value implication of their outputs is crucial. However, existing

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

EvoDRC: A Self-Evolving Agentic Framework for Automated DRC Violation Repair

DGX agent

arXiv:2607.20019v1 Announce Type: new Abstract: Design rule check (DRC) closure remains a major bottleneck in advanced-node physical design. Although detailed routers are rule-aware, residual design r

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance

DGX agent

arXiv:2607.19409v1 Announce Type: new Abstract: Recent advances in large language models have accelerated deployment of agentic systems in operational finance. Existing benchmarks emphasize measuring

model-releasesarxiv-cs-ai
23 Jul 2026
Tutorials

GraphContainer: A Unified Platform for Comparing and Debugging Graph RAG Methods

DGX agent

arXiv:2607.19362v1 Announce Type: new Abstract: Graph RAG mitigates hallucinations and stale knowledge in LLMs, particularly for multi-hop question answering. However, existing approaches remain highl

tutorialsarxiv-cs-ai
23 Jul 2026
Research

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose?

DGX agent

arXiv:2607.20284v1 Announce Type: new Abstract: The rapid development of multimodal large language models (MLLMs) has introduced a flexible paradigm for remote sensing image scene understanding (RSISU

researcharxiv-cs-cv
23 Jul 2026
Model Releases

NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs

DGX agent

arXiv:2607.14186v4 Announce Type: replace-cross Abstract: Scaling executable agent training data for LLM post-training is bottlenecked by substrate-bound methods that tie task generation to predefined

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Rethinking Uncertainty Evaluation in Large Language Models

DGX agent

arXiv:2607.19367v1 Announce Type: new Abstract: Calibration is the primary criterion for evaluating LLM confidence, but it is insufficient: it admits trivially incoherent estimators, depends on the ev

researcharxiv-cs-ai
23 Jul 2026
Hardware

Scaling Time Series Classification via XAI-Driven Data Reduction

DGX agent

arXiv:2607.15774v2 Announce Type: replace-cross Abstract: Explainable AI (XAI) for time series has seen significant algorithmic growth, but its utility in providing measurable performance gains for do

hardwarearxiv-cs-ai
23 Jul 2026
Model Releases

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

DGX agent

arXiv:2607.16741v2 Announce Type: replace Abstract: Burger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a mu

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Trusted Multi-View Deep Learning Classification of Fetal Congenital Heart Disease with Feature-level and Decision-level Fusion

DGX agent

arXiv:2606.15265v2 Announce Type: replace Abstract: Congenital heart disease (CHD) refers to the abnormal anatomical structure caused by the abnormal development of the heart and great vessels during

model-releasesarxiv-cs-cv
23 Jul 2026
Research

Understanding Developer Pain Points in Federated Learning: Insights from Stack Overflow and GitHub

DGX agent

arXiv:2607.19621v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training without centralizing raw data, but building and operating FL systems remains difficult du

researcharxiv-cs-ai
23 Jul 2026
← Previous
1…6162636465…109
Next →