AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Local Ai

Decentralized Multi-Agent Systems with Shared Context

DGX agent

arXiv:2606.10662v1 Announce Type: cross Abstract: Multi-agent systems (MAS) can scale large language model reasoning at test time by decomposing complex problems into parallel subtasks. However, most

local-aiarxiv-cs-ai
10 Jun 2026
Model Releases

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.11182v1 Announce Type: cross Abstract: In this paper, we propose EEVEE, the first multi-dataset test-time prompt learning framework for LLM agents, enabling test-time prompt learning under

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

DGX agent

arXiv:2606.10709v1 Announce Type: cross Abstract: The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Exploratory Responsiveness and Adaptive Rigidity under AI-Assisted Optimization

DGX agent

arXiv:2606.10086v1 Announce Type: new Abstract: This paper develops a theory of exploratory adaptation under AI-assisted optimization. The central argument is that the long-run adaptive effects of AI

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

DGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

safetyarxiv-cs-ai
10 Jun 2026
Research

Exploring the Design Space of Reward Backpropagation for Flow Matching

DGX agent

arXiv:2606.11075v1 Announce Type: new Abstract: Aligning text-to-image flow matching models with human preferences via direct reward backpropagation is sample-efficient but hampered by two well-known

researcharxiv-cs-lg
10 Jun 2026
Research

From Context-Aware to Conflict-Aware: Generalizing Contrastive Decoding for Knowledge Conflict in LLMs

DGX agent

arXiv:2606.10298v1 Announce Type: new Abstract: When large language models generate from retrieved or augmented contexts, conflicts between external context and parametric priors remain a central reli

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Frontier Coding Agents Use Metaprogramming to Adapt to Unfamiliar Programming Languages

DGX agent

arXiv:2606.10933v1 Announce Type: new Abstract: LLM-based coding agents are usually evaluated in familiar software settings: mainstream languages, common libraries, and public repositories. These benc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Gaming AI-Assisted Peer Reviews Poses New Risks to the Scientific Community

DGX agent

arXiv:2606.10159v1 Announce Type: cross Abstract: AI is increasingly used to support scientific peer review, from manuscript screening, reviewer assistance to editorial triage. Although such systems p

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines

DGX agent

arXiv:2606.09935v1 Announce Type: cross Abstract: AI-powered agents are increasingly embedded in continuous integration and continuous delivery/deployment (CI/CD) pipelines to autonomously review pull

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Google’s Gemini 3.5 Live Translate enables realistic translation at the speed of natural conversations

DGX agent

Google LLC’s newest artificial intelligence tool promises to bring real-time translation to every smartphone user, enabling more natural and fluid conversations between speakers of different languages

model-releasessiliconangle
10 Jun 2026
Model Releases

Janus: A Benchmark for Goal-Conditioned Information Distortion in LLMs

DGX agent

arXiv:2606.10852v1 Announce Type: cross Abstract: LLM deception is often evaluated through direct markers such as fabricated claims, explicit lies, or strategic concealment. However, many real-world m

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Less Context, More Accuracy: A Bi-Temporal Memory Engine for LLM Agents Where a Lean Retrieved Context Beats the Full History

DGX agent

arXiv:2606.09900v1 Announce Type: cross Abstract: Long-term memory is the missing layer for LLM agents: across sessions they forget, and the common workaround -- replaying the whole history into the p

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization

DGX agent

arXiv:2606.11180v1 Announce Type: new Abstract: Diffusion-based lip synchronization models achieve strong visual quality and audio-visual alignment, but full-sequence bidirectional attention and many

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

Magnetic HIP-NN for spin dynamics in disordered itinerant magnets

DGX agent

arXiv:2606.10349v1 Announce Type: cross Abstract: We present a magnetic extension of the Hierarchically Interacting Particle Neural Network (HIP-NN) that enables large-scale simulations of electron-me

model-releasesarxiv-cs-lg
10 Jun 2026
Research

Maximum Matching Accuracy: An Instance Segmentation Evaluation Metric Utilizing Globally Optimal Matching

DGX agent

arXiv:2606.10107v1 Announce Type: new Abstract: Reliable evaluation of instance segmentation models requires metrics that accurately and consistently reflect segmentation quality. However, the metrics

researcharxiv-cs-cv
10 Jun 2026
Model Releases

MoE Enhanced Federated Learning for Spatiotemporal Prediction

DGX agent

arXiv:2606.10499v1 Announce Type: cross Abstract: Traffic prediction is fundamental to intelligent transportation systems and urban computing, yet many cities continue to suffer from traffic data scar

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

MV-Actor: Aligning Multi-View Semantics and Spatial Awareness for Bimanual Manipulation

DGX agent

arXiv:2606.10899v1 Announce Type: new Abstract: Robotic manipulation has been widely applied in industrial scenarios. Compared with single-arm manipulation, bimanual manipulation is equipped with mult

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

Overcoming Rank Collapse in Feedback Alignment

DGX agent

arXiv:2606.11123v1 Announce Type: new Abstract: Backpropagation (BP) is widely viewed as biologically implausible, in part because it requires feedback weights to be the transpose of forward weights f

model-releasesarxiv-cs-lg
10 Jun 2026
Local Ai

PatchSTG: Scalable Spatiotemporal Graph Transformers for Traffic Forecasting on Irregular Sensor Networks

DGX agent

arXiv:2606.09872v1 Announce Type: cross Abstract: Traffic forecasting is a fundamental component of intelligent transportation systems, yet remains challenging in real-world settings due to irregular

local-aiarxiv-cs-ai
10 Jun 2026
Applications

PL-KKT-hPINN: Enforcing Nonlinear Equality Constraints on Neural Networks via Piecewise-Linear Projection

DGX agent

arXiv:2606.10682v1 Announce Type: new Abstract: While physics-informed neural networks (PINNs) have shown strong potential for process modeling, physical equations are only enforced as soft constraint

applicationsarxiv-cs-lg
10 Jun 2026
Model Releases

POPSICLE: Benchmark Datasets for Segmentation and Localization in CryoET

DGX agent

arXiv:2606.10255v1 Announce Type: cross Abstract: Cryo-electron tomography (cryoET) has emerged as a powerful tool in structural and cellular biology by enabling direct visualization of macromolecular

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

DGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

model-releasesarxiv-cs-cl
10 Jun 2026
Hardware

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

DGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

hardwarearxiv-cs-lg
10 Jun 2026
Model Releases

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

DGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

DGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

SocraticPO: Policy Optimization via Interactive Guidance

DGX agent

arXiv:2606.09887v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models usually supervises reasoning with scalar outcome rewards, such as binary correctness. Such rewar

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Sources: Microsoft is restricting employees from using Claude Fable 5 because of Anthropic's new 30-day data retention requirements (Tom Warren/The Verge)

DGX agent

Tom Warren / The Verge: Sources: Microsoft is restricting employees from using Claude Fable 5 because of Anthropic's new 30-day data retention requirements — Microsoft's legal teams are evaluating Ant

model-releasestechmeme
10 Jun 2026
Local Ai

Spatial-Omni: Spatial Audio Understanding Integration in Multimodal LLMs via FOA Encoding

DGX agent

arXiv:2606.10738v1 Announce Type: cross Abstract: Recent multimodal large language models mainly process audio as monaural signals, thereby discarding the spatial cues contained in spatial audio for s

local-aiarxiv-cs-ai
10 Jun 2026
Research

STORM: Stepwise Token Optimization with Reward-Guided Beam Search

DGX agent

arXiv:2606.10621v1 Announce Type: cross Abstract: Modern retrieval increasingly relies on dense and learned-sparse neural models that are effective but require encoding the entire corpus into a specia

researcharxiv-cs-ai
10 Jun 2026
Safety

Task Robustness via Re-Labelling Vision-Action Robot Data

DGX agent

arXiv:2606.10918v1 Announce Type: cross Abstract: The recent trend in scaling models for robot learning has resulted in impressive policies that can perform various manipulation tasks and generalize t

safetyarxiv-cs-lg
10 Jun 2026
Safety

TD-Grokking: Learning from Zero-Reward Problems by Training-Time Decomposition

DGX agent

arXiv:2606.09883v1 Announce Type: cross Abstract: Large language models (LLMs) have made remarkable progress in reasoning tasks, largely driven by post-training paradigms, especially reinforcement lea

safetyarxiv-cs-ai
10 Jun 2026
Agents

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

DGX agent

arXiv:2606.10747v1 Announce Type: new Abstract: As AI systems built from multiple language-model agents become more common, they are increasingly used to make decisions together: discussing, negotiati

agentsarxiv-cs-ai
10 Jun 2026
Tutorials

There is a lot of justified anger at Anthropic for sandbagging Fable 5 for AI development tasks. But an unanticipated side effect is that th…

DGX agent

There is a lot of justified anger at Anthropic for sandbagging Fable 5 for AI development tasks. But an unanticipated side effect is that third-party evaluators can no longer credibly use the model fo

tutorialsjeremy-howard--x
10 Jun 2026
Agents

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

DGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

UMI-Bench 1.0: An Open and Reproducible Real-World Benchmark for Tabletop Robotic Manipulation with UMI Data

DGX agent

arXiv:2606.10382v1 Announce Type: new Abstract: Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably outside curated demonstrations. This nee

model-releasesarxiv-cs-ro
10 Jun 2026
Research

Uncertainty-aware Multi-fidelity Closure via Conditional Normalizing Flows

DGX agent

arXiv:2606.09857v1 Announce Type: new Abstract: Reduced-order models (ROMs) provide an efficient surrogate for complex multiscale systems, but their predictive accuracy is often compromised by truncat

researcharxiv-cs-lg
10 Jun 2026
Safety

UniPET: a universal network for high-quality PET image denoising across varied dose reduction factors

DGX agent

arXiv:2606.11131v1 Announce Type: new Abstract: Most existing deep learning-based PET image denoising methods assume a fixed and known dose reduction factor (DRF) for low-dose PET images. However, the

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

What makes a harness a harness: necessary and sufficient conditions for an agent harness

DGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents

DGX agent

arXiv:2606.10267v1 Announce Type: cross Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot manipulation, by using high-level VLM plan

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice

DGX agent

arXiv:2606.10249v1 Announce Type: new Abstract: We examine whether graph neural network (GNN) design rules generalize across benchmark families by studying aggregator selection (sum, mean, max) on 24

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

WHU-Infra3D: A Full-stack Multi-modal Dataset and Benchmark for 3D Roadside Infrastructure Inventory

DGX agent

arXiv:2606.09882v1 Announce Type: new Abstract: The paradigm of digital twin cities is shifting from coarse visual mapping toward more precise and actionable digitization of urban assets. However, exi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

With SoftBank apparently struggling to get a margin loan against it’s OpenAI shares, it’s maybe time to repost this, from two years ago. The…

DGX agent

With SoftBank apparently struggling to get a margin loan against it’s OpenAI shares, it’s maybe time to repost this, from two years ago. The vast majority of my earlier worries remain: 9 reasons that

model-releasesgary-marcus--x
10 Jun 2026
Tutorials

A Systematic Study of Behavioral Cloning for Scientific Data Annotation

DGX agent

arXiv:2606.07568v1 Announce Type: cross Abstract: Scientific data annotation, such as tracking animals in video or proofreading neural reconstructions, remains bottlenecked by the 'last mile' problem:

tutorialsarxiv-cs-ai
9 Jun 2026
Tutorials

Adversarial Attack and Disturbance Detection by Hadamard-Coded Output Representations for Object Detection and Semantic Segmentation

DGX agent

arXiv:2606.09536v1 Announce Type: new Abstract: Conventional one-hot encodings often yield poorly calibrated models, being overconfident under attack, and letting entropy-based detection algorithms fa

tutorialsarxiv-cs-cv
9 Jun 2026
Local Ai

AGENTSERVESIM: A Hardware-aware Simulator for Multi-Turn LLM Agent Serving

DGX agent

arXiv:2606.09613v1 Announce Type: cross Abstract: Multi-turn LLM agents interleave model calls with external tool invocations, shifting serving from stateless request processing to stateful program ex

local-aiarxiv-cs-ai
9 Jun 2026
Model Releases

AgroOmni: A Large-Scale Multi-view Agricultural Dataset for Cross-Scale Multimodal Reasoning

DGX agent

arXiv:2603.14342v2 Announce Type: replace-cross Abstract: Modern agricultural data is sourced from diverse platforms and spans multiple spatial scales, ranging from ground-level close-up photography t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity (Anthropic)

DGX agent

Anthropic: Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity — Today we're launching Claude

model-releasestechmeme
9 Jun 2026
← Previous
1…687688689690691…1359
Next →