AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
Human
88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
9 Jun 2026

Supracompetitive Pricing Under AI Monoculture

Model ReleasesDGX agent

arXiv:2601.01279v3 Announce Type: replace-cross Abstract: When competing sellers delegate pricing to a shared AI model, such as a large language model, correlated recommendations combined with perform

SurfDesign: Effective Protein Design on Molecular Surfaces

Model ReleasesDGX agent

arXiv:2606.07567v1 Announce Type: cross Abstract: Protein function is largely determined by molecular surface geometry and physicochemical complementarity, yet most protein design methods condition on

Sustainability and Artificial Intelligence: Necessary, Challenging, and Promising Intersections

TutorialsDGX agent

arXiv:2606.09006v1 Announce Type: cross Abstract: Both digital economy and digital technology researchers increasingly recognize the need to better address the role that artificial intelligence (AI) p

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SVRG and Beyond via Posterior Correction

ResearchDGX agent

arXiv:2512.01930v2 Announce Type: replace-cross Abstract: Stochastic Variance Reduced Gradient (SVRG) and its variants aim to speed-up training by using gradient corrections. Originally proposed over

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

Model ReleasesDGX agent

arXiv:2606.07682v1 Announce Type: cross Abstract: AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex env

SwiftVR: Real-Time One-Step Generative Video Restoration

HardwareDGX agent

arXiv:2606.09516v1 Announce Type: new Abstract: Real-time video restoration (VR) for live streams requires high-resolution outputs under strict per-frame latency constraints. Existing one-step diffusi

Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Topics, and Models

SafetyDGX agent

arXiv:2606.08451v1 Announce Type: cross Abstract: Safety-aligned large language models often exhibit sycophancy, which is the tendency to affirm users' opinions regardless of factual accuracy. Althoug

Syll: Open-Source Personal Automation with Cross-Surface Execution

AgentsDGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

SafetyDGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

Symskill: Symbol and Skill Co-Invention for Data-Efficient and Reactive Long-Horizon Manipulation

ResearchDGX agent

arXiv:2510.01661v3 Announce Type: replace Abstract: Multi-step manipulation in dynamic environments remains challenging. Imitation learning (IL) is reactive but lacks compositional generalization, sin

SynManDex: Synthesizing Human-like Dexterous Grasps from Synthetic Human Pre-Grasps

ResearchDGX agent

arXiv:2606.09798v1 Announce Type: new Abstract: Human hand-object interactions encode functional intent, but direct transfer to robotic hands often fails under morphology, contact, and reachability co

Synthetic but Not Realistic: The Evaluation Challenge in Generative Modelling for Structured Electronic Medical Records

ApplicationsDGX agent

arXiv:2606.08903v1 Announce Type: new Abstract: Synthetic healthcare data are widely proposed as privacy-preserving substitutes for real patient data, yet their evaluation remains dominated by statist

SynthICL: Scalable In-context Imitation Learning with Synthetic Data

SafetyDGX agent

arXiv:2606.08154v1 Announce Type: new Abstract: In-context imitation learning (ICIL) enables robots to learn new tasks from a small number of demonstrations by conditioning a pre-trained policy on tas

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

Model ReleasesDGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

Systems-Level Planning and Coordination of Truck-Drone Collaborative Delivery Networks

SafetyDGX agent

arXiv:2606.08738v1 Announce Type: cross Abstract: Urban last-mile parcel delivery increasingly relies on heterogeneous fleets whose performance depends on timely coordination, reliable communication,

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

Model ReleasesDGX agent

arXiv:2606.09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

Model ReleasesDGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

Taming Perception Jitter: Uncertainty-Aware LiDAR Object Detection for Reliable Motion Classification

AgentsDGX agent

arXiv:2606.09350v1 Announce Type: cross Abstract: Reliable motion classification is critical for autonomous driving, as false dynamic predictions of static objects can cascade into unnecessary planner

TAMUNA: Doubly Accelerated Distributed Optimization under Partial Participation

Local AiDGX agent

arXiv:2302.09832v4 Announce Type: replace Abstract: In distributed optimization and federated learning, slow and costly communication between parallel devices and the central server constitutes the pr

TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks

HardwareDGX agent

arXiv:2510.16028v4 Announce Type: replace-cross Abstract: Neural networks increasingly run on hardware outside the user's control (cloud GPUs, inference marketplaces). Yet ML-as-a-Service reveals litt

Targeting World Models to Compromise Robot Learning Pipelines

SafetyDGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

TBD-VLA: Temporal Block Diffusion Vision Language Action Model

ApplicationsDGX agent

arXiv:2606.07895v1 Announce Type: new Abstract: Discrete Vision-Language-Action (VLA) models typically formulate action generation as next-token prediction over discretized action spaces, conditioning

Teacher-Free Self-Training Amplifies but Does Not Compound: A Pass@K Crossover on a Free-Verifier Domain

Local AiDGX agent

arXiv:2606.07856v1 Announce Type: new Abstract: When a language model trains on its own verified outputs, does it acquire capability beyond its base, or merely get better at expressing capability the

TeamHerald@CHIPSAL 2026: Hate Speech Detection and Sentiment Analysis of Nepali Memes using Transformer-based Architectures and Ensemble Learning

ResearchDGX agent

arXiv:2606.08770v1 Announce Type: cross Abstract: The analysis of internet memes in the Nepali language is complicated by frequent code-mixing and a lack of established baseline resources. While memes

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

Temporal-Aware Reasoning Optimization for Video Temporal Grounding

Local AiDGX agent

arXiv:2606.09248v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) have achieved remarkable progress in video temporal grounding with reinforcement learning for generating reaso

Temporal Coverage over Density: Parsimonious Training-Set Design for ML Climate Downscaling

ResearchDGX agent

arXiv:2606.07898v1 Announce Type: new Abstract: High-resolution regional climate simulations provide critical information for climate impacts assessments but remain computationally expensive, motivati

Tensorizing Engram: Sharing Latents Across N-Gram Embeddings is Beneficial in LLMs

ResearchDGX agent

arXiv:2606.08347v1 Announce Type: cross Abstract: Modern language models represent text using discrete token-level embeddings, which forces recurring multi-token patterns to be learned implicitly acro

Test-Time Adaptive Composition for Machine Learning as a Service (MLaaS) in IoT Environments

ResearchDGX agent

arXiv:2606.07685v1 Announce Type: cross Abstract: The dynamic nature of Internet of Things (IoT) environments affects the long-term effectiveness of Machine Learning as a Service (MLaaS) compositions.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning

ResearchDGX agent

arXiv:2606.08231v1 Announce Type: new Abstract: Test-time Scaling (TTS) has emerged as a pivotal research direction for enhancing model performance by dynamically allocating computational resources du

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

SafetyDGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

SafetyDGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

Model ReleasesDGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

The CIFAR Synthetic Evidence Corpus for Detecting AI-Generated Evidence

TutorialsDGX agent

arXiv:2606.07916v1 Announce Type: new Abstract: The growing ability of generative models to produce realistic documents poses a direct challenge to evidentiary workflows in the justice system and the

The Confidence Trap: Calibration Attacks for Graph Neural Networks

SafetyDGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

The Cross-Architecture Substrate: A Domain-Transcendent, Calibration-Surviving Geometric Invariant of Modern Vision Encoders

SafetyDGX agent

arXiv:2606.07882v1 Announce Type: cross Abstract: Different vision neural networks -- trained to classify, contrast, reconstruct, or match images to text -- should have correspondingly different inter

The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.07950v1 Announce Type: new Abstract: RL with verifiable rewards can substantially improve LLM reasoning, yet standard GRPO-style training often treats easy, hard, and learnable questions al

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

SafetyDGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective Default Lock-In

SafetyDGX agent

arXiv:2606.08172v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate high-stakes interactions in finance, medicine, and mental-health support, yet users have limited con

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

SafetyDGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

Model ReleasesDGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

The Label Horizon Paradox: Rethinking Supervision Targets in Financial Forecasting

ResearchDGX agent

arXiv:2602.03395v4 Announce Type: replace Abstract: While deep learning has revolutionized financial forecasting through sophisticated architectures, the design of the supervision signal itself is rar

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.07861v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) excel at multimodal understanding and reasoning, yet their fine-grained visual perception remains underexplored.

The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes

ResearchDGX agent

arXiv:2402.08922v3 Announce Type: replace Abstract: Large-scale black-box models have become ubiquitous across numerous applications. Understanding the influence of individual training data sources on

The Montparnasse Algorithm for RNA Design

Model ReleasesDGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

The Need for Neural ISP in the Small-Pixel Era: How Shrinking Pixels Push Optics to the Limit and Neural Restoration Pushes Back

Local AiDGX agent

arXiv:2606.07675v1 Announce Type: cross Abstract: Smartphone telephoto cameras are approaching a 'telephoto physics wall': as pixel pitches shrink toward sub-0.5 micron, the optics remain limited by g

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers

TutorialsDGX agent

arXiv:2606.07587v1 Announce Type: new Abstract: LLM routing has become a popular approach to improve the cost-quality trade-off of LLM services by dynamically selecting a model for each query. Recent

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

Model ReleasesDGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

The Spectral Dynamics and Noise Geometry of Muon

SafetyDGX agent

arXiv:2606.08388v1 Announce Type: new Abstract: Muon replaces a matrix gradient G=USigma V^op by its polar factor UV^op. This keeps the singular directions selected by the gradient, but makes the upda

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

AgentsDGX agent

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

The Topological Dual of a Dataset: A Logic-to-Topology Encoding for AlphaGeometry-Style Data

ResearchDGX agent

arXiv:2604.18050v2 Announce Type: replace Abstract: AlphaGeometry represents a milestone in neuro-symbolic reasoning, yet its architecture faces a log-linear scaling bottleneck within its symbolic ded

The Value of Personalized Recommendations: Evidence from Netflix

ResearchDGX agent

arXiv:2511.07280v5 Announce Type: replace-cross Abstract: Personalized recommendation systems shape much of user choice online, yet their targeted nature makes separating out the value of recommendati

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

Model ReleasesDGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

Model ReleasesDGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

Think Before You Act: Intention-Guided Reasoning for LLM-Based Location Prediction

SafetyDGX agent

arXiv:2606.08122v1 Announce Type: new Abstract: Predicting a user's next Point-of-Interest (POI) based on their historical check-in records is a fundamental task in location-based services. While rece

Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.04805v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have attracted much attention due to their exceptional performance. However, their performance mainly stems from think

Thinking Without Images: Internalizing Visual Manipulation with On-Policy Self-Distillation

Local AiDGX agent

arXiv:2606.08719v1 Announce Type: new Abstract: ''Thinking with Images'' has emerged as an effective paradigm for fine-grained visual reasoning: by explicitly zooming into relevant regions and reasoni

Thresholded Local Hyper-Flow Diffusion

ResearchDGX agent

arXiv:2606.09340v1 Announce Type: new Abstract: Local Hyper-Flow Diffusion (HFD) gives an edge-size-independent Cheeger-type guarantee for seeded clustering in general submodular hypergraphs, but exis

TianJi-Environ: An Autonomous AI Scientist for Atmospheric Environmental Research

AgentsDGX agent

arXiv:2606.07697v1 Announce Type: cross Abstract: As atmospheric environmental prediction continues to improve, interpretable validation of pollution mechanisms and feedback processes has become a mai

TIDE: Task-Isolated Diffusion for Unified Video Editing and Generation

ResearchDGX agent

arXiv:2606.08260v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers have driven rapid progress in video generation and editing, yet these capabilities are still handled by separa

← Previous
1…455456457458459…1049
Next →