AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Safety

Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Topics, and Models

DGX agent

arXiv:2606.08451v1 Announce Type: cross Abstract: Safety-aligned large language models often exhibit sycophancy, which is the tendency to affirm users' opinions regardless of factual accuracy. Althoug

safetyarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Syll: Open-Source Personal Automation with Cross-Surface Execution

DGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

agentsarxiv-cs-ai
9 Jun 2026
Safety

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

DGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

safetyarxiv-cs-ai
9 Jun 2026
Research

Symskill: Symbol and Skill Co-Invention for Data-Efficient and Reactive Long-Horizon Manipulation

DGX agent

arXiv:2510.01661v3 Announce Type: replace Abstract: Multi-step manipulation in dynamic environments remains challenging. Imitation learning (IL) is reactive but lacks compositional generalization, sin

researcharxiv-cs-ro
9 Jun 2026
Research

SynManDex: Synthesizing Human-like Dexterous Grasps from Synthetic Human Pre-Grasps

DGX agent

arXiv:2606.09798v1 Announce Type: new Abstract: Human hand-object interactions encode functional intent, but direct transfer to robotic hands often fails under morphology, contact, and reachability co

researcharxiv-cs-ro
9 Jun 2026
Applications

Synthetic but Not Realistic: The Evaluation Challenge in Generative Modelling for Structured Electronic Medical Records

DGX agent

arXiv:2606.08903v1 Announce Type: new Abstract: Synthetic healthcare data are widely proposed as privacy-preserving substitutes for real patient data, yet their evaluation remains dominated by statist

applicationsarxiv-cs-lg
9 Jun 2026
Safety

SynthICL: Scalable In-context Imitation Learning with Synthetic Data

DGX agent

arXiv:2606.08154v1 Announce Type: new Abstract: In-context imitation learning (ICIL) enables robots to learn new tasks from a small number of demonstrations by conditioning a pre-trained policy on tas

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

DGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Systems-Level Planning and Coordination of Truck-Drone Collaborative Delivery Networks

DGX agent

arXiv:2606.08738v1 Announce Type: cross Abstract: Urban last-mile parcel delivery increasingly relies on heterogeneous fleets whose performance depends on timely coordination, reliable communication,

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

DGX agent

arXiv:2606.09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

DGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Taming Perception Jitter: Uncertainty-Aware LiDAR Object Detection for Reliable Motion Classification

DGX agent

arXiv:2606.09350v1 Announce Type: cross Abstract: Reliable motion classification is critical for autonomous driving, as false dynamic predictions of static objects can cascade into unnecessary planner

agentsarxiv-cs-cv
9 Jun 2026
Local Ai

TAMUNA: Doubly Accelerated Distributed Optimization under Partial Participation

DGX agent

arXiv:2302.09832v4 Announce Type: replace Abstract: In distributed optimization and federated learning, slow and costly communication between parallel devices and the central server constitutes the pr

local-aiarxiv-cs-lg
9 Jun 2026
Hardware

TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks

DGX agent

arXiv:2510.16028v4 Announce Type: replace-cross Abstract: Neural networks increasingly run on hardware outside the user's control (cloud GPUs, inference marketplaces). Yet ML-as-a-Service reveals litt

hardwarearxiv-cs-ai
9 Jun 2026
Safety

Targeting World Models to Compromise Robot Learning Pipelines

DGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

safetyarxiv-cs-ai
9 Jun 2026
Applications

TBD-VLA: Temporal Block Diffusion Vision Language Action Model

DGX agent

arXiv:2606.07895v1 Announce Type: new Abstract: Discrete Vision-Language-Action (VLA) models typically formulate action generation as next-token prediction over discretized action spaces, conditioning

applicationsarxiv-cs-cv
9 Jun 2026
Local Ai

Teacher-Free Self-Training Amplifies but Does Not Compound: A Pass@K Crossover on a Free-Verifier Domain

DGX agent

arXiv:2606.07856v1 Announce Type: new Abstract: When a language model trains on its own verified outputs, does it acquire capability beyond its base, or merely get better at expressing capability the

local-aiarxiv-cs-lg
9 Jun 2026
Research

TeamHerald@CHIPSAL 2026: Hate Speech Detection and Sentiment Analysis of Nepali Memes using Transformer-based Architectures and Ensemble Learning

DGX agent

arXiv:2606.08770v1 Announce Type: cross Abstract: The analysis of internet memes in the Nepali language is complicated by frequent code-mixing and a lack of established baseline resources. While memes

researcharxiv-cs-ai
9 Jun 2026
Model Releases

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

DGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Temporal-Aware Reasoning Optimization for Video Temporal Grounding

DGX agent

arXiv:2606.09248v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) have achieved remarkable progress in video temporal grounding with reinforcement learning for generating reaso

local-aiarxiv-cs-cv
9 Jun 2026
Research

Temporal Coverage over Density: Parsimonious Training-Set Design for ML Climate Downscaling

DGX agent

arXiv:2606.07898v1 Announce Type: new Abstract: High-resolution regional climate simulations provide critical information for climate impacts assessments but remain computationally expensive, motivati

researcharxiv-cs-lg
9 Jun 2026
Research

Tensorizing Engram: Sharing Latents Across N-Gram Embeddings is Beneficial in LLMs

DGX agent

arXiv:2606.08347v1 Announce Type: cross Abstract: Modern language models represent text using discrete token-level embeddings, which forces recurring multi-token patterns to be learned implicitly acro

researcharxiv-cs-lg
9 Jun 2026
Research

Test-Time Adaptive Composition for Machine Learning as a Service (MLaaS) in IoT Environments

DGX agent

arXiv:2606.07685v1 Announce Type: cross Abstract: The dynamic nature of Internet of Things (IoT) environments affects the long-term effectiveness of Machine Learning as a Service (MLaaS) compositions.

researcharxiv-cs-ai
9 Jun 2026
Research

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning

DGX agent

arXiv:2606.08231v1 Announce Type: new Abstract: Test-time Scaling (TTS) has emerged as a pivotal research direction for enhancing model performance by dynamically allocating computational resources du

researcharxiv-cs-cv
9 Jun 2026
Safety

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

DGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

safetyarxiv-cs-ai
9 Jun 2026
Safety

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

DGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

DGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

model-releasesarxiv-cs-ai
9 Jun 2026
Tutorials

The CIFAR Synthetic Evidence Corpus for Detecting AI-Generated Evidence

DGX agent

arXiv:2606.07916v1 Announce Type: new Abstract: The growing ability of generative models to produce realistic documents poses a direct challenge to evidentiary workflows in the justice system and the

tutorialsarxiv-cs-ai
9 Jun 2026
Safety

The Confidence Trap: Calibration Attacks for Graph Neural Networks

DGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Cross-Architecture Substrate: A Domain-Transcendent, Calibration-Surviving Geometric Invariant of Modern Vision Encoders

DGX agent

arXiv:2606.07882v1 Announce Type: cross Abstract: Different vision neural networks -- trained to classify, contrast, reconstruct, or match images to text -- should have correspondingly different inter

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.07950v1 Announce Type: new Abstract: RL with verifiable rewards can substantially improve LLM reasoning, yet standard GRPO-style training often treats easy, hard, and learnable questions al

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

DGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective Default Lock-In

DGX agent

arXiv:2606.08172v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate high-stakes interactions in finance, medicine, and mental-health support, yet users have limited con

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

DGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

DGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

model-releasesarxiv-cs-lg
9 Jun 2026
Research

The Label Horizon Paradox: Rethinking Supervision Targets in Financial Forecasting

DGX agent

arXiv:2602.03395v4 Announce Type: replace Abstract: While deep learning has revolutionized financial forecasting through sophisticated architectures, the design of the supervision signal itself is rar

researcharxiv-cs-lg
9 Jun 2026
Model Releases

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models

DGX agent

arXiv:2606.07861v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) excel at multimodal understanding and reasoning, yet their fine-grained visual perception remains underexplored.

model-releasesarxiv-cs-ai
9 Jun 2026
Research

The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes

DGX agent

arXiv:2402.08922v3 Announce Type: replace Abstract: Large-scale black-box models have become ubiquitous across numerous applications. Understanding the influence of individual training data sources on

researcharxiv-cs-lg
9 Jun 2026
Model Releases

The Montparnasse Algorithm for RNA Design

DGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

The Need for Neural ISP in the Small-Pixel Era: How Shrinking Pixels Push Optics to the Limit and Neural Restoration Pushes Back

DGX agent

arXiv:2606.07675v1 Announce Type: cross Abstract: Smartphone telephoto cameras are approaching a 'telephoto physics wall': as pixel pitches shrink toward sub-0.5 micron, the optics remain limited by g

local-aiarxiv-cs-cv
9 Jun 2026
Tutorials

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers

DGX agent

arXiv:2606.07587v1 Announce Type: new Abstract: LLM routing has become a popular approach to improve the cost-quality trade-off of LLM services by dynamically selecting a model for each query. Recent

tutorialsarxiv-cs-lg
9 Jun 2026
Model Releases

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

DGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

The Spectral Dynamics and Noise Geometry of Muon

DGX agent

arXiv:2606.08388v1 Announce Type: new Abstract: Muon replaces a matrix gradient G=USigma V^op by its polar factor UV^op. This keeps the singular directions selected by the gradient, but makes the upda

safetyarxiv-cs-lg
9 Jun 2026
Agents

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

DGX agent

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

agentsarxiv-cs-ai
9 Jun 2026
Research

The Topological Dual of a Dataset: A Logic-to-Topology Encoding for AlphaGeometry-Style Data

DGX agent

arXiv:2604.18050v2 Announce Type: replace Abstract: AlphaGeometry represents a milestone in neuro-symbolic reasoning, yet its architecture faces a log-linear scaling bottleneck within its symbolic ded

researcharxiv-cs-ai
9 Jun 2026
Research

The Value of Personalized Recommendations: Evidence from Netflix

DGX agent

arXiv:2511.07280v5 Announce Type: replace-cross Abstract: Personalized recommendation systems shape much of user choice online, yet their targeted nature makes separating out the value of recommendati

researcharxiv-cs-lg
9 Jun 2026
Model Releases

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

DGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

DGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

model-releasesarxiv-cs-lg
9 Jun 2026
← Previous
1…602603604605606…1344
Next →