AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
Human
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
24 Jul 2026

Synthetic minority data is redundant or invalid: a data-dependent validity theory and a de-biased test

SafetyDGX agent

arXiv:2607.20787v1 Announce Type: cross Abstract: For two decades, the standard remedy for class-imbalanced learning has been to fabricate synthetic minority examples, and the standard evidence of the

T-STAR: A Large-Scale Benchmark for Spatio-Temporal Panoptic Scene Graph Generation in Satellite Video

Model ReleasesDGX agent

arXiv:2607.21228v1 Announce Type: new Abstract: Structured understanding of satellite video is essential for advancing dynamic geospatial scene analysis from low-level perception to high-level cogniti

TableVerse: A Large-scale Tabletop Dataset with Real-world Grounded Layouts for Generalizable Manipulation

ApplicationsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.21017v1 Announce Type: new Abstract: The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While

TabPFN beyond Tabular Data: Calibration and Accuracy on Multimodal Embeddings

ResearchDGX agent

arXiv:2607.11007v2 Announce Type: replace Abstract: Few-shot multimodal classification commonly attaches a lightweight head, such as k-nearest neighbors, logistic regression, or a linear SVM, to a fro

Tackling Heterogeneity in Federated Learning via Variance-Reduced Boltzmann Sampling within Homogeneous Social Coalitions

ResearchDGX agent

arXiv:2506.02897v3 Announce Type: replace Abstract: Federated Learning (FL) enables privacy-preserving collaborative model training, but its effectiveness is often limited by client data heterogeneity

TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework

AgentsDGX agent

arXiv:2511.05385v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) utilizes external knowledge to augment Large Language Models' (LLMs) reliability. For flexibility, agenti

Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain

Model ReleasesDGX agent

arXiv:2607.20510v1 Announce Type: new Abstract: We introduce Telco-GAIA, a bilingual, multi-modal benchmark for evaluating tool-using agents on the data of a real-world telecommunications operator. Te

Tencent WorkBuddy Bench: A Multi-Domain Coding-Agent Benchmark with Contamination-Resistant Task Construction

Model ReleasesDGX agent

arXiv:2607.20911v1 Announce Type: new Abstract: We introduce Tencent WorkBuddy Bench, a multi-domain evaluation suite for coding agents; this report documents its construction methodology, scoring pro

Test-Time Scaling via Error Localization

Local AiDGX agent

arXiv:2607.21453v1 Announce Type: new Abstract: Scaling inference-time computation has emerged as a reliable method to improve the performance of large language models on complex reasoning and program

Texture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion Model

ResearchDGX agent

arXiv:2607.21504v1 Announce Type: new Abstract: Numerous 3D assets are discarded due to low texture resolution, while current super-resolution models ignore texture maps and focus on natural images. A

thaulab@EEUCA 2026: Who Said What to Whom? A Targeting-Aware Neural-Symbolic Pipeline for Gaming Toxicity Detection

SafetyDGX agent

arXiv:2607.20447v1 Announce Type: new Abstract: This paper describes our system for the EEUCA 2026 Shared Task on toxicity classification in gaming chat. We implement a three-stage pipeline combining

The Active Ingredient in Muon's Grokking

ResearchDGX agent

arXiv:2607.20512v1 Announce Type: cross Abstract: The Muon optimizer reaches the grokking threshold on modular arithmetic faster than AdamW. Prior work attributes this to 'spectral-norm constraints pl

The Boundaries of Automation: A Theory of Persistent Human Participation

ApplicationsDGX agent

arXiv:2607.21547v1 Announce Type: new Abstract: The rapid progress of AI has intensified the long-standing pursuit of automation: replacing human participation with algorithms wherever possible. Impli

The Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually Works

SafetyDGX agent

arXiv:2607.21273v1 Announce Type: new Abstract: Dense per-step supervision is an appealing remedy for sparse-reward, long-horizon LLM agents: reward the agent for predicting its next observation, and

The Devil is in the Spectrum: Mitigating Representation Collapse in LLMs via Topologically Regularized Side-Path

Model ReleasesDGX agent

arXiv:2607.20484v1 Announce Type: new Abstract: Large Language Models (LLMs) are fundamentally limited by representation collapse, a bottleneck that severely degrades long-context performance. We iden

The Geometry of Personality: Activation Steering with Jungian Cognitive Functions

Model ReleasesDGX agent

arXiv:2607.20803v1 Announce Type: cross Abstract: Activation steering enables control and interpretation of LLMs, yet existing work primarily models personality through static trait frameworks such as

The Hidden Footprint: Making Storage a First-Class Metric for LLM Agent Evaluation

Model ReleasesDGX agent

arXiv:2607.11149v3 Announce Type: replace Abstract: LLM agent benchmarks measure task completion, reliability, and inference cost, but not the persistent data an agent run leaves on disk, including lo

The Human-AI Substitution Principle: When will you be replaced by AI in your organization?

ResearchDGX agent

arXiv:2607.20781v1 Announce Type: new Abstract: Artificial Intelligence (AI) is rapidly transforming organizations, raising a fundamental organizational and economic question: when will a human employ

The Price of Hidden Curvature: An widetilde{Omega} (d^{5/4} sqrt{T}) Lower Bound for Bandit Convex Optimization

ResearchDGX agent

arXiv:2607.18652v2 Announce Type: replace-cross Abstract: We establish a widetildeOmega(d^{5/4}sqrt T) lower bound on the minimax expected regret of stochastic bandit convex optimization of 1-Lipschit

The RealDefocus Benchmark for Defocus Deblurring

Model ReleasesDGX agent

arXiv:2607.21078v1 Announce Type: new Abstract: Single-Image Defocus Deblurring (SIDD) aims to recover an all-in-focus image from a single defocused observation, but rigorous and reproducible evaluati

The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

Model ReleasesDGX agent

arXiv:2607.21118v1 Announce Type: new Abstract: This paper presents a review of the second LoViF Challenge on Real-World All-in-One Image Restoration. The challenge aims to advance unified image resto

The Storyteller in the Model: Narrative Pattern Inheritance, Escalation Dynamics, and Alignment Governance in LLMs

SafetyDGX agent

arXiv:2607.20449v1 Announce Type: cross Abstract: LLMs are trained predominantly on human-authored text, yet the structural and narrative conventions embedded in that text are rarely examined as a sou

The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning

ResearchDGX agent

arXiv:2607.20952v1 Announce Type: cross Abstract: Latent, or silent, reasoning lets language models carry out intermediate computation in continuous vector space instead of words, and is widely assume

Thermodynamic Weight Decay: Exploring Grokking Acceleration via Attention Specific Heat

ResearchDGX agent

arXiv:2607.20552v1 Announce Type: new Abstract: Grokking -- the delayed generalization of neural networks long after they have memorized their training data -- wastes thousands of training epochs and

Thinkink: 2D Spatial Ink-native Interaction with LLMs

ResearchDGX agent

arXiv:2607.21468v1 Announce Type: cross Abstract: People often use handwritten notes and sketches to externalize ideas for ideation. To integrate large language models (LLMs) into this practice, we pr

THOR: A Theta-Gamma Hierarchical Oscillatory Reasoning Framework for Multi-hop QA

ResearchDGX agent

arXiv:2607.20459v1 Announce Type: cross Abstract: Multi-hop question answering requires retrieving and integrating evidence from multiple contexts. Despite the rapid progress of current research, mult

Three-Pronged Spectral Control for Federated Parameter Efficient Fine Tuning

Model ReleasesDGX agent

arXiv:2607.20914v1 Announce Type: new Abstract: Federated parameter-efficient fine-tuning (PEFT) enables communication-efficient adaptation of large pretrained models on decentralized edge data, but i

Through-the-Earth Magnetic Induction Communication and Networking: A Comprehensive Survey

ResearchDGX agent

arXiv:2510.14854v4 Announce Type: cross Abstract: Magnetic induction (MI) communication (MIC) has emerged as a promising candidate for underground communication networks due to its excellent penetrati

Token Budget Saturation and Mechanistic Early Detection of Reasoning Non-Convergence in Chain-of-Thought Models

Model ReleasesDGX agent

arXiv:2607.21433v1 Announce Type: cross Abstract: Chain-of-thought reasoning models such as DeepSeek-R1-Distill-Qwen-7B exhibit a bimodal convergence pattern: generations either terminate within a tok

Token-Level Entropy Reveals Demographic Disparities in Large Language Models

SafetyDGX agent

arXiv:2501.19337v5 Announce Type: replace Abstract: A name alone measurably reshapes a language model's next-token distribution before a single token is sampled. We measure full-vocabulary Shannon ent

TopoGuard: Graph Theory Based Defenses Against Split-Knowledge Attacks on RAG

ApplicationsDGX agent

arXiv:2607.20437v1 Announce Type: new Abstract: Production Retrieval Augmented Generation (RAG) systems rely on aggregating multiple external documents to answer complex queries. However, the retrieve

TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics

Model ReleasesDGX agent

arXiv:2602.19313v2 Announce Type: replace-cross Abstract: General-purpose robot learning requires dense, instruction-conditioned feedback that can distinguish meaningful task progress from stalled, fa

torchsom: The Reference PyTorch Library for Self-Organizing Maps

Model ReleasesDGX agent

arXiv:2510.11147v2 Announce Type: replace-cross Abstract: This paper introduces torchsom, an open-source Python library that provides a reference implementation of the Self-Organizing Map (SOM) in PyT

TOUR: A Trajectory-Level Unlearning Benchmark for Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.21111v1 Announce Type: cross Abstract: Offline Reinforcement Learning (RL) agents are trained on fixed behavioral trajectories, which makes trajectory-level deletion important when selected

Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry

Local AiDGX agent

arXiv:2607.21495v1 Announce Type: new Abstract: AI agents are increasingly created inside organizations by non-engineering users through low-code, no-code, and conversational development environments.

Toward cryptographically verifiable authorization for autonomous AI agents: A security hypothesis, preliminary formal model, and proof-of-concept implementation

SafetyDGX agent

arXiv:2607.21325v1 Announce Type: cross Abstract: Autonomous AI agents increasingly execute actions, invoke tools, and operate on protected resources with limited human oversight. Existing authenticat

Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.21496v1 Announce Type: cross Abstract: Cognitive impairment (CI) is a growing public health concern. Early and accurate diagnosis is critical for enabling timely intervention and improving

Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry

Model ReleasesDGX agent

arXiv:2607.20778v1 Announce Type: new Abstract: Weather forecasting foundation models (FMs) are increasingly fine-tuned to predict air quality, offering fast global pollution forecasts at lower comput

Towards a Certifying Grounder

ResearchDGX agent

arXiv:2607.21199v1 Announce Type: cross Abstract: Grounding, the translation of high-level theories into equivalent quantifier-free formulas, is a crucial step in declarative solving, yet it has so fa

Towards an Automated Test of LLM Security Knowledge

Model ReleasesDGX agent

arXiv:2607.18496v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for a range of software, hardware and human-centered security tasks. Consequently, LLM perf

Towards Capability-Aware Traversability Navigation for Unstructured Environments

ResearchDGX agent

arXiv:2607.20679v1 Announce Type: new Abstract: Estimating traversability in unstructured environments requires conditioning on robot embodiment, as the same terrain can be traversable for one platfor

Towards Faithful Graph Explanations with Synergistic Edge Effects via Granular Balls

Model ReleasesDGX agent

arXiv:2607.21381v1 Announce Type: new Abstract: Instance-level explanations aim to reveal the rationale behind a model's decisions for a specific graph. Previous methods explain graph neural networks

Towards Privacy-Preserving Federated Prompt Tuning under Data Heterogeneity: A Subspace-Decomposed Expert Approach

Local AiDGX agent

arXiv:2607.21417v1 Announce Type: new Abstract: Federated prompt tuning (FPT) enables collaborative adaptation of vision--language models (VLMs) using lightweight prompts. Existing methods often addre

Towards Robust Iris Recognition Through Occlusion Identification and Conditional Diffusion-Based Reconstruction

ResearchDGX agent

arXiv:2607.21545v1 Announce Type: new Abstract: Iris recognition is a reliable biometric approach that identifies individuals using the distinctive and stable texture of the iris. However, recognition

Traceable Scholarship: Page Anchors and Ariadne's Thread for Humanistic Inquiry in the Age of Generative AI

AgentsDGX agent

arXiv:2607.20916v1 Announce Type: new Abstract: Generative AI lets large language models produce scholarly-looking text within seconds, yet fluency does not equal valid explanation. The deepest risk i

Tractable Hierarchical Control of Autoregressive Language Models

ResearchDGX agent

arXiv:2607.20483v1 Announce Type: new Abstract: Constraining the generation of autoregressive large language models (LLMs) is an important component of integrating language models into formal systems.

Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers

HardwareDGX agent

arXiv:2512.16615v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) set the state of the art in visual generation, yet their quadratic self-attention cost fundamentally limits scaling to

Training Large Language Models for Self-Explanation Faithfulness

Model ReleasesDGX agent

arXiv:2607.21090v1 Announce Type: cross Abstract: We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated r

TransBiolab: A Real-World Multi-View Dataset of Cluttered Transparent Biomedical Objects

Model ReleasesDGX agent

arXiv:2607.21071v1 Announce Type: new Abstract: Autonomous biomedical laboratories increasingly rely on visual perception to recognize, localize, and manipulate transparent plasticware, yet high-quali

Transformer-Assisted LLM-Based Source Code Summarisation: to Enable More Secure Software Development

ResearchDGX agent

arXiv:2607.20933v1 Announce Type: cross Abstract: Neural Source Code Summarisation (NSCS) aims to generate natural language summaries of source code to improve developers' and maintainers' understandi

Transformer-based Diffusion models for Hydrological Time Series Probabilistic Imputation and Forecasting

ResearchDGX agent

arXiv:2607.21200v1 Announce Type: cross Abstract: The modeling of hydrometeorological time series with limited observations is a key challenge in the monitoring of hydro-systems and water resources, a

Transition-Related Potentials as Markers of Narrative Comprehension in Continuous EEG

ResearchDGX agent

arXiv:2607.20720v1 Announce Type: cross Abstract: Harnessing the potential of electroencephalography (EEG) for brain research is fundamentally limited by intrinsic noise and the diffuse projection of

TwistedMerge: Certified Higher-Order Diagnostics and Abstention for Model Merging

SafetyDGX agent

arXiv:2607.20887v1 Announce Type: cross Abstract: Model merging combines independently trained or fine-tuned models, but pairwise alignability does not imply globally consistent alignment. We formulat

U-CFR: Uncertainty-Guided Cascade Forward Refinement for Interactive Segmentation

Model ReleasesDGX agent

arXiv:2607.20705v1 Announce Type: cross Abstract: Interactive image segmentation is critical for efficient image annotation; however, existing methods often require many corrective clicks or rely on p

Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement

ResearchDGX agent

arXiv:2607.20529v1 Announce Type: cross Abstract: Large Language Model (LLM) ensembles are increasingly used to improve reliability by combining predictions from multiple LLMs. However, existing aggre

UnDA: Unpaired Domain Alignment for Cross-Modal Knowledge Transfer in Medical Imaging

SafetyDGX agent

arXiv:2607.21546v1 Announce Type: new Abstract: Multimodal based approaches often outperform single modality approaches in downstream tasks as the different modalities provide complementary informatio

Understanding Critical Thinking in Generative Artificial Intelligence Use: Development, Validation, and Correlates of the Critical Thinking in AI Use Scale

SafetyDGX agent

arXiv:2512.12413v2 Announce Type: replace Abstract: Generative AI tools are increasingly embedded in everyday work and learning, yet their fluency, opacity, and propensity to hallucinate mean that use

Unified Video Dense Prediction from Disjoint Data

ResearchDGX agent

arXiv:2607.21592v1 Announce Type: new Abstract: Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragment

Unlearning Under Imbalance: Benchmarking Fairness in Multimodal LLM Unlearning

Model ReleasesDGX agent

arXiv:2607.21300v1 Announce Type: cross Abstract: Machine unlearning has emerged as a tool for removing personal data from trained models to comply with recent AI regulations. To evaluate unlearning e

Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Ghana

ResearchDGX agent

arXiv:2607.21559v1 Announce Type: new Abstract: A consensus anomaly detection framework was applied to monthly malaria surveillance data from Ghana (2014-2023) to identify atypical transmission patter

← Previous
1…171172173174175…998
Next →