AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
30 Jun 2026

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

Model ReleasesDGX agent

arXiv:2606.29467v1 Announce Type: new Abstract: Medical question-answering benchmarks rarely cover the maternal, neonatal, child, and reproductive-health questions a nurse-midwife asks, and, to our kn

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

AgentsDGX agent

arXiv:2606.30296v1 Announce Type: new Abstract: Multi-round reflection lets agents built on large language models recover from failures within a single task, but each task remains an isolated episode:

MARS: A neurosymbolic approach for interpretable drug discovery

SafetyDGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MAVIN: Multi-Shot Audio-Visual Generation with Narrative Control

AgentsDGX agent

arXiv:2606.29473v1 Announce Type: new Abstract: While recent generative models produce high-fidelity videos, they struggle with the complex narrative control required for coherent multi-shot audio-vis

MemLeak: Diagnosing Information Leaks in Multimodal Agent Memory

Model ReleasesDGX agent

arXiv:2606.29788v1 Announce Type: new Abstract: When a multimodal AI agent is asked to forget a fact, current memory systems usually delete the text entry and report success. We find that the fact can

Meta-learning as a principle for human-like visual representations

SafetyDGX agent

arXiv:2606.28399v1 Announce Type: new Abstract: The structure of human visual representations underpins our capacity for adaptive behaviour. While pretrained neural networks model human visual represe

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

SafetyDGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning

SafetyDGX agent

arXiv:2602.13562v2 Announce Type: replace-cross Abstract: While reasoning models have achieved remarkable success in complex reasoning tasks, their increasing power necessitates stringent safety measu

Modernizing financial services with deployment freedom and transformational AI with AlloyDB Omni

Local AiDGX agent

The financial services industry (FSI) operates under a unique set of non-negotiable requirements: the need for strict regulatory compliance, sub-millisecond transactional speeds, and security that ver

Monte Carlo Energy Aggregation for Mobile 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2606.30017v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting have demonstrated unprecedented success in novel view synthesis. However, the substantial inference and storage

Multi-Agent Route Planning as a QUBO Problem

Model ReleasesDGX agent

arXiv:2602.07913v2 Announce Type: replace Abstract: Multi-Agent Route Planning considers selecting vehicles, each associated with a single predefined route, such that route-level coverage utility is m

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

SafetyDGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

Multiply Robust Causal Mediation Analysis with Continuous Treatments

Model ReleasesDGX agent

arXiv:2105.09254v4 Announce Type: replace-cross Abstract: In many applications, researchers are interested in the direct and indirect causal effects of a treatment or exposure on an outcome of interes

Muon learns balanced solutions in matrix factorization without slow saddle-to-saddle dynamics

Model ReleasesDGX agent

arXiv:2606.30509v1 Announce Type: new Abstract: Matrix factorization (i.e., problems of the form min_{mathbf{P},mathbf{Q}} |mathbf{M}^star - mathbf{P}^opmathbf{Q}|_F^2) is a minimal learning problem t

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, wr…

Model ReleasesDGX agent

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, write a post, give 1 or 2 talks on it, rewrite the post, give

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

AgentsDGX agent

arXiv:2606.29824v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel as static solvers, transforming them into autonomous agents remains challenging. This transition requires con

// Neural procedural memory // Good paper on agent memory beyond prompt retrieval. NPM stores procedural skills as activation steering vecto…

Model ReleasesDGX agent

// Neural procedural memory // Good paper on agent memory beyond prompt retrieval. NPM stores procedural skills as activation steering vectors distilled from contrastive historical experience. Textual

Neural Subspace Reallocation: Continual Learning as Retrieval-Based Subspace Memory Management

Model ReleasesDGX agent

arXiv:2606.30067v1 Announce Type: cross Abstract: We introduce Neural Subspace Reallocation (NSR), which reframes continual learning as memory management over parameter subspaces. Instead of treating

Notes (and a Pelican) on Claude Sonnet 5 - the new tokenizer makes it ~1.4x more expensive for English, ~1.33x more expensive for Spanish bu…

Model ReleasesDGX agent

Notes (and a Pelican) on Claude Sonnet 5 - the new tokenizer makes it ~1.4x more expensive for English, ~1.33x more expensive for Spanish but roughly the same price for Simplified Mandarin https://sim

On FrontierCode (Extended), our benchmark for real-world engineering tasks that grades mergeability and quality, Sonnet 5 scores 53.8% and h…

Model ReleasesDGX agent

On FrontierCode (Extended), our benchmark for real-world engineering tasks that grades mergeability and quality, Sonnet 5 scores 53.8% and has a 57.6% pass rate (higher than Opus 4.8). Note: These rel

On the Necessity of a Liquid Substrate for Mesh Intelligence

AgentsDGX agent

arXiv:2606.28413v1 Announce Type: cross Abstract: A mesh of sovereign agents has no center: no shared clock, no shared model, and no coordinator to gather data or retrain. Its competence rests on each

On the Policy Gradient Foundations of Group Relative Policy Optimization: Credit Assignment, Gradient Sparsity, and Rank Collapse

Model ReleasesDGX agent

arXiv:2606.29238v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) eliminates the learned critic in PPO by using the mean reward of grouped rollouts as a baseline. We provide a

Online Data Selection for Instruction Tuning via Gaussian Processes

Local AiDGX agent

arXiv:2606.30077v1 Announce Type: cross Abstract: With Large Language Model (LLM) pre-training and fine-tuning shifting its focus from data volume to data quality, quality data selection has emerged a

Optimization Dynamics Imprint Semantic Specificity in Contrastive Embedding Norms

ResearchDGX agent

arXiv:2606.30625v1 Announce Type: cross Abstract: Contrastive embedding models trained with scale-invariant losses are typically paired with distance metrics like cosine similarity, effectively ignori

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

Model ReleasesDGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

PCGD: Physics-Guided Conditional Graph Diffusion for TCAD Device Simulation

Model ReleasesDGX agent

arXiv:2606.29272v1 Announce Type: new Abstract: Technology computer-aided design (TCAD) semiconductor device simulation is fundamentally constrained by the high computational cost of iteratively solvi

Perforce launches Agentic Gateway to govern AI agents and cut token costs

Model ReleasesDGX agent

Perforce Software Inc. today expanded its Perforce Intelligence lineup with an agentic gateway for managing artificial intelligence agents, an autonomous testing platform driven by natural language an

Personalizing MLLMs via Reinforced Multimodal Reference Game

ResearchDGX agent

arXiv:2606.28845v1 Announce Type: new Abstract: Personalizing Multimodal Large Language Models (MLLMs) aims to recognize users' unique concepts from visual data and provide personalized responses. Alt

Pie launches with $19.5M to bring AI marketing to small businesses

Model ReleasesDGX agent

Pie Tech Inc., a startup using artificial intelligence to provide growth tools for small businesses, today officially launched with an announcement that it has raised 19.5 million in new funding to ex

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

Model ReleasesDGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

Pooled Leaderboards Hide System-Specific Winners: A Reporting-Protocol Audit of Offline Root-Cause Analysis Benchmarks

Model ReleasesDGX agent

arXiv:2606.29159v1 Announce Type: new Abstract: Offline root-cause-analysis (RCA) benchmarks commonly rank methods by a single pooled top-1 accuracy across multiple subsystems, and engineers often rea

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

Model ReleasesDGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

Post-training for Efficient Communication via Convention Formation

Model ReleasesDGX agent

arXiv:2508.06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions. In contra

Projected Exploitability Descent for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

Model ReleasesDGX agent

arXiv:2606.29169v1 Announce Type: cross Abstract: Many important games have more than two players and imperfect information. Existing approaches for computing Nash equilibrium, the central game-theore

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

Model ReleasesDGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

Qwen publishes new work on RL coding agents. (bookmark it) The idea is to continually build a verification system that co-evolves with AI ag…

Model ReleasesDGX agent

Qwen publishes new work on RL coding agents. (bookmark it) The idea is to continually build a verification system that co-evolves with AI agents. LLMs suffer from all sorts of reward hacking issues. T

Randomized neural operator for parametric PDEs with fast training and conformal uncertainty quantification

Model ReleasesDGX agent

arXiv:2606.29440v1 Announce Type: new Abstract: Repeatedly solving parametric PDEs is essential for uncertainty quantification, design optimization and inverse problems, but conventional neural operat

Reachability Guarantees for Cart-Pole Swing-Up and Stabilization

Model ReleasesDGX agent

arXiv:2606.28627v1 Announce Type: cross Abstract: The cart-pole swing-up is a canonical benchmark for nonlinear control of underactuated systems, yet an end-to-end guarantee linking the global swing-u

Read the full Aston Martin F1 team interview with @aidangomez here: https://www.astonmartinf1.com/en-GB/news/feature/perspectives-aidan-gome…

Model ReleasesDGX agent

This post links to a full interview with Aidan Gomez conducted by the Aston Martin F1 team, published on their official website under their 'Perspectives' feature section. The interview likely covers

Recursive Self-Evolving Agents via Held-Out Selection

Model ReleasesDGX agent

arXiv:2606.28374v1 Announce Type: new Abstract: LLM agents are increasingly improved without weight updates by evolving a natural-language artifact, such as reflections, workflows, playbooks, cheatshe

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

Model ReleasesDGX agent

arXiv:2606.30294v1 Announce Type: new Abstract: Live product demonstrations are a recurring, high-cost activity in software organizations: a human presenter must select features, dispatch the correspo

Relevance Is Not Permission: Warranted Attention for Value Contributions

Local AiDGX agent

arXiv:2606.30139v1 Announce Type: new Abstract: Relevance is not permission. Attention lets a model read key-value items related to the current query, but it does not guarantee that the value contribu

Research Entity Extraction and Topic Detection from UKRI Grant Proposals

Model ReleasesDGX agent

arXiv:2606.30304v1 Announce Type: cross Abstract: This paper presents preliminary findings from a UKRI-funded Metascience project comparing three LLM-based approaches, GPT-4o, Mistral, and a bespoke a

Residual-Guided Dictionary Learning for Spectrally Accurate Koopman Approximation

Model ReleasesDGX agent

arXiv:2606.29083v1 Announce Type: cross Abstract: Koopman theory promises linear structure in nonlinear dynamics, but numerical Koopman spectra are easy to compute and hard to trust. A finite EDMD mat

Resolution Thresholds in VLM Detection of Harmful ASCII Art Across Construction Modes and Languages

SafetyDGX agent

arXiv:2606.29649v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) are increasingly deployed as content moderation tools, yet they remain vulnerable to jailbreak attacks in which harm

Rethinking Forgery Attacks on Semantic Watermarks in Black-Box Settings: A Geometric Distortion Perspective

ResearchDGX agent

arXiv:2606.29807v1 Announce Type: cross Abstract: Recent studies have shown that semantic watermarks, which embed information into the initial noise of latent diffusion models (LDMs), are vulnerable t

Reweighting Framewise Attention in Video Transformers for Facial Expression Understanding

ResearchDGX agent

arXiv:2606.30611v1 Announce Type: new Abstract: Understanding facial expressions in videos requires modeling subtle and localized facial dynamics under unconstrained conditions. Although recent Vision

RoAd-RL: A Unified Library and Benchmark for Robust Adversarial Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.29867v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in robotics and autonomous systems, yet remains vulnerable to adversarial perturbat

SA-Homo: Scale Adaptive Homography Estimation for Scale Variation Scenarios

Model ReleasesDGX agent

arXiv:2606.30408v1 Announce Type: new Abstract: Homography estimation, as one of the fundamental problems in computer vision, remains challenged by scale variation scenarios where image pairs potentia

Scalar Representations of Neural Network Training Dynamics

Model ReleasesDGX agent

arXiv:2606.30384v1 Announce Type: new Abstract: Training in artificial neural networks can be viewed as a trajectory evolving through a high-dimensional loss landscape. However, the large number of tr

Scaling Textual Gradients via Sampling-Based Momentum

Model ReleasesDGX agent

arXiv:2506.00400v4 Announce Type: replace-cross Abstract: LLM-based prompt optimization, which uses LLM-provided ``textual gradients'' (feedback) to refine prompts, has emerged as an effective method

Seed-to-Seed: Unpaired Image Translation in Diffusion Seed Space

ResearchDGX agent

arXiv:2409.00654v2 Announce Type: replace Abstract: We introduce Seed-to-Seed Translation (StS), a novel approach that combines GANs and diffusion models (DMs) for unpaired Image-to-Image Translation.

Selective Memory Retention for Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2606.29178v1 Announce Type: new Abstract: When does retention matter for memory-augmented LLM agents? We study this with TraceRetain, a lightweight framework for bounded external memory in froze

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Fir…

Model ReleasesDGX agent

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Firewall. And more! ✨ All shipped in June! 🤯 And there's a good

Sequential Hiring of Contingent Workers Through Learning-Based Optimization

Model ReleasesDGX agent

arXiv:2606.18438v2 Announce Type: replace-cross Abstract: In this paper, we study a sequential workforce management problem in a contingent labor setting with uncertainty in both worker production and

Sequential Planning via Anchored Robotic Keypoints

Model ReleasesDGX agent

arXiv:2606.30613v1 Announce Type: new Abstract: We present Sequential Planning via Anchored Robotic Keypoints, SPARK, a training-free neurosymbolic manipulation system that reaches 43.7% on six LIBERO

SIGNET: Motion-Level Knowledge Transfer for Cross-Language Sign Language Translation

ResearchDGX agent

arXiv:2606.28626v1 Announce Type: new Abstract: Sign language translation (SLT) remains challenging due to its high spatio-temporal complexity, long sequences, and the need to model multiple articulat

SIR: Structured Image Representations for Explainable Robot Learning

SafetyDGX agent

arXiv:2606.30101v1 Announce Type: cross Abstract: Existing robot policies based on learned visual embeddings lack explicit structure and are sensitive to visual distractions. Thus, the representations

Situation Perception: A Necessary Primitive to Artificial Superintelligence

ResearchDGX agent

arXiv:2606.30481v1 Announce Type: cross Abstract: Current large language models are extraordinary statistical engines. They compress vast amounts of text into useful patterns and can explain science,

SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning

ResearchDGX agent

arXiv:2602.02472v2 Announce Type: replace-cross Abstract: Progressive Learning (PL) reduces pre-training computational overhead by gradually increasing model scale. While prior work has extensively ex

← Previous
1…636637638639640…1053
Next →