AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
1 Jul 2026

{Phi}eat: Physically Grounded Material Feature Representation

ResearchDGX agent

arXiv:2511.11270v2 Announce Type: replace Abstract: While foundation models have emerged as general-purpose visual backbones, their representations are primarily optimized for semantics and lack expli

PPT-Eval: A Benchmark for Computer-Use Agents on PowerPoint Tasks

Model ReleasesDGX agent

arXiv:2606.31154v1 Announce Type: cross Abstract: Creating and editing slides is a rich, multimodal activity that is ubiquitous in professional and educational settings, making it an ideal testbed for

PrISM-IQA: Image Quality Assessment Made Practical for Smartphone Photography

Model ReleasesDGX agent

arXiv:2606.31626v1 Announce Type: new Abstract: Existing smartphone image quality assessment (IQA) methods commonly reduce perceptual quality to a single score. However, this scalar formulation is poo

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Prompting Robot Teams with Natural Language

SafetyDGX agent

arXiv:2509.24575v2 Announce Type: replace-cross Abstract: This paper presents a framework to prompt multi-robot teams with high-level tasks using natural language expressions. Our objective is to use

Qualified Educational Capacity Planning under Heterogeneous Student Support Needs: A Synthetic Benchmark and Decision-Support Framework

Model ReleasesDGX agent

arXiv:2606.30650v1 Announce Type: cross Abstract: Educational support services often face a qualified-capacity problem: staff time is scarce, qualifications decay, new support needs can appear before

Quality-Aware Modulation for Diffusion Transformers

ResearchDGX agent

arXiv:2606.30934v1 Announce Type: new Abstract: Modern text-to-image diffusion models, such as diffusion transformers (DiT), rely on timestep or prompt embeddings to modulate the strength of the denoi

Review Residuals: Update-Conditioned Residual Gating for Transformers

Model ReleasesDGX agent

arXiv:2606.31859v1 Announce Type: cross Abstract: Residual connections add every sublayer's proposed update with a fixed coefficient of one; the network never evaluates whether an update is reliable b

Sequential sparse Gaussian process quantile regression

Model ReleasesDGX agent

arXiv:2606.31284v1 Announce Type: new Abstract: Quantile regression aims to estimate the conditional quantiles of a response variable from observed data. In a Bayesian setting, Gaussian process quanti

SkillSpotter: Pose-Aware Multi-View Skilled Action Detection and Grading in Ego-Exo Videos

Model ReleasesDGX agent

arXiv:2606.31127v1 Announce Type: cross Abstract: To enable personalized, real-time coaching using Augmented Reality glasses or fixed camera setups in domains such as sports, cooking, or music, a syst

Some of LangSmith's biggest and most engaged customers trace with: - @aisdk - Claude Agents SDK - OpenAI Agent SDK - No framework at all And…

Model ReleasesDGX agent

Some of LangSmith's biggest and most engaged customers trace with: - @aisdk - Claude Agents SDK - OpenAI Agent SDK - No framework at all And I'm sure we'll see tons of http://pi.dev in the near future

TabPATE: Differentially Private Tabular In-Context Learning Without Public Data

ResearchDGX agent

arXiv:2606.31474v1 Announce Type: new Abstract: Tabular foundation models enable accurate in-context learning (ICL) from small labeled datasets, but the private records placed in context can leak thro

TaxoMIL: Taxonomy-Constrained Learning for Hierarchical Whole Slide Image Analysis

Model ReleasesDGX agent

arXiv:2606.31100v1 Announce Type: new Abstract: Whole slide image (WSI) analysis is central to computational pathology, with multiple instance learning (MIL) emerging as the standard pipeline for slid

TerraDiT-Omega: Unified Spatial Control for Satellite Image Synthesis with Any Geospatial Primitive

Local AiDGX agent

arXiv:2606.31029v1 Announce Type: new Abstract: Generative models have achieved remarkable progress, yet applying them to satellite imagery remains challenging. Unlike natural imagery, satellite scene

The automation rate of remote projects has increased ~4x in the past five months.

Model ReleasesDGX agent

The automation rate of remote projects has increased ~4x in the past five months. New Remote Labor Index results: AI automation of real remote work is increasing fast. Claude Fable 5 now completes 16.

The discussion here on AI futures can be a little too credulous of company visions. People tend to push what they have. The three big AI lab…

ApplicationsDGX agent

The discussion here on AI futures can be a little too credulous of company visions. People tend to push what they have. The three big AI labs will say bigger models are the future. Every other firm ha

The latest AI news we announced in June 2026

Model ReleasesDGX agent

Google's June 2026 AI announcements included the launch of Gemini 3.5 Live Translate, the latest features in Android 17 and the new Google Home Speaker built for Gemini. The updates introduced new fea

Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning

SafetyDGX agent

arXiv:2606.31599v1 Announce Type: cross Abstract: Vision-language models (VLMs) combining reinforcement learning (RL) ignite remarkable progress in multimodal reasoning, yet still struggle with medica

Tone-Conditioned Curriculum Learning for Low-Resource Bantu Speech Recognition

ApplicationsDGX agent

arXiv:2606.31642v1 Announce Type: new Abstract: Southern Bantu languages are spoken by over 80 million people, yet current foundation ASR models still produce zero-shot WER above 100%, which limits pr

Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

Model ReleasesDGX agent

arXiv:2606.30755v1 Announce Type: cross Abstract: Claw-like AI agents (e.g., OpenClaw) are always-on processes with persistent access to credentials, files, tools, and external services. They take on

Unveiling Transferability in Trajectory Prediction via Latent Scene Embeddings

AgentsDGX agent

arXiv:2606.30777v1 Announce Type: new Abstract: The growing availability of trajectory datasets has fueled major advances in data-driven motion prediction. Yet, models trained on one dataset often fai

Wait, am I Being Fair? Characterizing Deductive Stereotyping and Mitigating It with Fair-GCG

SafetyDGX agent

arXiv:2606.30989v1 Announce Type: cross Abstract: Warning: This paper contains several toxic and offensive statements. While reasoning generally improves fairness in recent large language models (LLMs

WarpI2I: Image Warping for Image-to-Image Translation

ResearchDGX agent

arXiv:2606.31018v1 Announce Type: new Abstract: Image-to-image (I2I) translation has achieved strong results in tasks like human relighting and driving scene translation using latent diffusion models

What If We Allocate Test-Time Compute Adaptively?

Model ReleasesDGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineering

Model ReleasesDGX agent

arXiv:2606.30911v1 Announce Type: new Abstract: ML engineering agents waste compute rediscovering known techniques because every competition is a cold start. We present HASTE, a hierarchical multi-age

30 Jun 2026

A Comparative Study on Affective Cues in Text Embeddings Across Psychological Emotion Theories

Model ReleasesDGX agent

arXiv:2606.29068v1 Announce Type: cross Abstract: Text encoders are known for their utility in natural language processing, as they are able to efficiently compress inputs into dense vectors while pre

A Stochastic--Geometric Theory of Scaling Laws in Grokking

Model ReleasesDGX agent

arXiv:2606.30388v1 Announce Type: cross Abstract: Delayed generalization (ie~grokking) refers to the phenomenon in which a neural network fits its training data early in training but only begins to ge

Agile Reinforcement Learning through Separable Neural Architecture and Applications

Model ReleasesDGX agent

arXiv:2601.23225v2 Announce Type: replace-cross Abstract: Deep reinforcement learning (RL) is increasingly deployed in resource-constrained environments, yet go-to function approximators - multilayer

Analysis of Adam Algorithms for Stochastic Dynamic Systems

Model ReleasesDGX agent

arXiv:2606.28879v1 Announce Type: new Abstract: The adaptive moment estimation algorithm, known as Adam, is widely used in modern machine learning, owing to its low per-iteration complexity and strong

Anthropic launches Claude Sonnet 5, saying it nears Opus 4.8 performance at lower prices and is substantially better than Sonnet 4.6 for agentic work (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic launches Claude Sonnet 5, saying it nears Opus 4.8 performance at lower prices and is substantially better than Sonnet 4.6 for agentic work — Claude Sonnet 5 is built to be the mo

BabyHuBERT: Multilingual Self-Supervised Learning for Segmenting Speakers in Child-Centered Long-Form Recordings

ResearchDGX agent

arXiv:2509.15001v3 Announce Type: replace-cross Abstract: Child-centered daylong recordings are essential for studying early language development, but existing speech models trained on clean adult dat

Bayesian Best-Arm Identification with Abstention: A Polynomial-to-Exponential Phase Transition

Model ReleasesDGX agent

arXiv:2606.29203v1 Announce Type: new Abstract: We study the Bayesian fixed-budget best-arm identification problem in which a learner can abstain from making a terminal recommendation. Subject to an a

Benchmark AUC Is Not Deployable Reliability: A Cross-Dataset Audit of Off-the-Shelf Features for Surveillance Video Anomaly Detection

Model ReleasesDGX agent

arXiv:2606.29506v1 Announce Type: new Abstract: Automated 'suspicious behavior' flagging is a headline promise of AI surveillance, and the field reports high frame-level ROC-AUC on standard video anom

Bilevel Optimization for Neural Architecture Search

ResearchDGX agent

arXiv:2606.29582v1 Announce Type: cross Abstract: Bilevel optimization has become an influential and widely adopted framework for addressing hierarchical optimization problems in machine learning, pro

Bridging the NISQ and Fault-Tolerant Regimes: Generative-ML-Assisted Quantum Selected CI for Molecular Simulations

Model ReleasesDGX agent

arXiv:2606.30551v1 Announce Type: cross Abstract: Calculation of binding energies for protein-ligand molecular systems requires accurate treatment of the electronic structure, a quantum chemistry prob

Claude Sonnet 5 costs 2 per 1M input tokens and 10 per 1M output tokens through August 31, after which prices rise to 3 and 15, respectively (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: Claude Sonnet 5 costs 2 per 1M input tokens and 10 per 1M output tokens through August 31, after which prices rise to 3 and 15, respectively — Anthropic is upgrading Claude Sonnet,

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation

SafetyDGX agent

arXiv:2606.29805v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are prone to hallucination as their generation preferences are insufficiently calibrated to visual evidence, ca

Clinical Reasoning Graphs: Structured Evaluation of LLM Diagnostic Reasoning Reveals Competence Without Consistency

ResearchDGX agent

arXiv:2606.29876v1 Announce Type: cross Abstract: Modern large language models (LLMs) reach 60-70% diagnostic accuracy on complex clinical case benchmarks, but accuracy alone cannot distinguish stable

CoGS: Compositional Dynamic Human-Object Scenes Gaussian Splatting from Monocular Video

Model ReleasesDGX agent

arXiv:2606.28820v1 Announce Type: new Abstract: Reconstructing dynamic human--object interaction scenes from monocular video is difficult because the human, manipulated object, and background obey dif

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

Model ReleasesDGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

Constrained Tabular Diffusion for Finance

ApplicationsDGX agent

arXiv:2606.28674v1 Announce Type: cross Abstract: Generative models in finance face the dual challenge of producing realistic data while satisfying strict regulatory and economic objectives, a require

Contrastive vision-language learning with paraphrasing and negation

TutorialsDGX agent

arXiv:2511.16527v2 Announce Type: replace Abstract: Contrastive vision-language models continue to be the dominant approach for image-text retrieval. Contrastive Language-Image Pre-training (CLIP) tra

Decision-Value Attribution in Predict-then-Optimize Systems

TutorialsDGX agent

arXiv:2606.29878v1 Announce Type: new Abstract: Predictive models are increasingly embedded in operational decision-making, yet standard explanation methods typically explain forecasts rather than the

Decomposing Memorization Reduction in Privacy-Preserving Fine-Tuning of SLMs for CSIRTs

ResearchDGX agent

arXiv:2606.28479v1 Announce Type: cross Abstract: CSIRTs increasingly fine tune language models on vulnerability scan records, but these records expose internal network topology and create privacy ris

Diagnosing and Repairing Factual Errors in RAG under Budget Constraints

Local AiDGX agent

arXiv:2606.29377v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) improves the factuality of large language models by grounding responses in external evidence, yet real-world deploy

Diffusion Fine-tuning with Rewarded Moment Matching Distillation

SafetyDGX agent

arXiv:2606.30414v1 Announce Type: new Abstract: Distillation and Reinforcement Learning (RL) fine-tuning are the primary pillars of diffusion post-training. While traditionally studied in isolation, t

Digitizing Coaching Intelligence: An Agentic Framework for Holistic Athlete Profiling using VLM and RAG

Model ReleasesDGX agent

arXiv:2606.28570v1 Announce Type: cross Abstract: Athlete assessment is a critical process for tracking physical progress and identifying elite talent. However, during mass recruitment drives, traditi

Discovering Collaboration from Novelty: Random Network Distillation for Clustered Federated Learning

Local AiDGX agent

arXiv:2606.30499v1 Announce Type: new Abstract: Federated Learning often suffers under non-independently and identically distributed data, where a single global model may fail to represent the diversi

Dockerless: Environment-Free Program Verifier for Coding Agents

Model ReleasesDGX agent

arXiv:2606.28436v1 Announce Type: cross Abstract: Program verifiers play a central role in training coding agents, including selecting trajectories for supervised fine-tuning (SFT) and providing rewar

DriftGuard: Safety-Aware Multi-Monitor Detection and Selective Adaptation for Evolving Toxicity Moderation

Local AiDGX agent

arXiv:2606.28725v1 Announce Type: new Abstract: Automated toxicity moderation systems operate in dynamic online environments where harmful behavior evolves through coded language, shifting targets, an

Edit in 2D, Verify in 3D: Reinforcement Learning for Multi-view Consistent Scene Editing

ApplicationsDGX agent

arXiv:2603.03143v2 Announce Type: replace-cross Abstract: Leveraging the priors of 2D diffusion models for 3D editing has emerged as a promising paradigm. However, multi-view consistency remains chall

Efficient Visual Pointing for Embodied AI:Agent-Driven Data Synthesis, Cross-Block Attention, and Iterative Correction

Model ReleasesDGX agent

arXiv:2606.29850v1 Announce Type: new Abstract: Visual pointing maps a language instruction to pixel co ordinates, a core skill for embodied AI. We describe our PointArena 2026 solution, which achieve

Entropy-Gated Latent Recursion

ResearchDGX agent

arXiv:2606.16620v2 Announce Type: replace-cross Abstract: Inference-time scaling has become the dominant lever for improving language-model reasoning, but existing methods derive rollout diversity fro

EVAF: A Test-Retest Protocol for Selective Parametric Consolidation

Model ReleasesDGX agent

arXiv:2606.29916v1 Announce Type: cross Abstract: Long-running language agents need mechanisms for deciding which experiences should persist after the working context is gone. Retrieval systems can re

Every article saying 'THIS IS THE NEW JOB OF THE AI ERA' is one of two jobs. And the first is a lie imo. The first job, is usually something…

Model ReleasesDGX agent

Every article saying 'THIS IS THE NEW JOB OF THE AI ERA' is one of two jobs. And the first is a lie imo. The first job, is usually something that an AI lab is hiring for and the news blows it complete

Explainability-Aware Frustum Attack: Exposing Structural Vulnerabilities in LiDAR-Based 3D Object Detectors

Model ReleasesDGX agent

arXiv:2606.29963v1 Announce Type: new Abstract: The structural vulnerabilities of point cloud-based 3D object detectors remain poorly understood. Prior work has studied adversarial robustness primaril

Exploiting Local Flatness for Efficient Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2606.29952v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) data is crucial for reliable machine learning deployment. Among detection strategies, post-hoc methods are particu

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks

ApplicationsDGX agent

arXiv:2606.30452v1 Announce Type: new Abstract: Tabular data dominate the landscape of data science, increasingly attracting innovative machine learning models and tailored benchmarks. Yet, little is

Few-Shot Domain Incremental Learning via Continual Vision-Language Consolidation

Model ReleasesDGX agent

arXiv:2606.30190v1 Announce Type: cross Abstract: Existing domain-incremental learning (DIL) strategies call for massive amounts of data to adapt to new domains and suffer from the overfitting problem

From Accuracy to Visual Dependence: Auditing and Filtering Modality Collapse in Traffic VideoQA

Model ReleasesDGX agent

arXiv:2606.30220v1 Announce Type: new Abstract: High benchmark accuracy does not guarantee genuine use of visual evidence. We study this problem in traffic accident Video Question Answering (VideoQA),

Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing

Model ReleasesDGX agent

arXiv:2606.30599v1 Announce Type: new Abstract: Existing instruction-based video editing datasets commonly focus on single-task appearance editing, failing to meet the complex creative demands of real

← Previous
1…526527528529530…1061
Next →