AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
9 Jun 2026

Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

Model ReleasesDGX agent

arXiv:2606.08340v1 Announce Type: new Abstract: As language models are increasingly deployed as autonomous agents, they must coordinate with others over long horizons in open-ended interactive tasks.

Beyond Consistency: Preserving Temporal Structure in Zero-Shot Video Editing

Model ReleasesDGX agent

arXiv:2606.08780v1 Announce Type: new Abstract: Existing zero-shot video editing methods rely on pre-trained diffusion models, successfully achieving spatial control and basic temporal consistency but

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (and Fail) Red Team Attacks

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.07833v1 Announce Type: cross Abstract: Standard AI red teaming evaluations reduce adversarial campaigns to a single binary outcome, attack success rate (ASR), not taking into account the se

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

Model ReleasesDGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning

ResearchDGX agent

arXiv:2606.09393v1 Announce Type: new Abstract: Image and video captioning are fundamental tasks that bridge the visual and linguistic domains, playing a critical role in pre-training Large Vision-Lan

CHROMA: Detecting AI-Generated Images through Inter-Channel Color-Space Correlations

Model ReleasesDGX agent

arXiv:2606.08864v1 Announce Type: new Abstract: The rapid adoption of diffusion and large-scale generative models has made it increasingly challenging to distinguish synthetic imagery from real photog

Claude Fable 5 and new AI safety fables

Model ReleasesDGX agent

This article discusses Claude Fable 5, likely exploring Anthropic's latest version of their AI model and examining new fables or narratives related to AI safety concepts. The piece probably analyzes h

Claude Fable 5 is now available on Databricks, fully governed through Unity AI Gateway

Model ReleasesDGX agent

Claude Fable 5 is now available as a model option on the Databricks platform, integrated with Unity AI Gateway to provide governance controls for enterprise users. This integration allows organization

Claude Mythos went from “too dangerous to release” to publicly available (with some extra guard rails) in two months. And y’all fell for Ant…

Model ReleasesDGX agent

Gary Marcus critiques Anthropic's rapid shift in positioning Claude from a model deemed too dangerous for public release to one made widely available with safety measures, suggesting this represents i

ComplexConstraints and Beyond: Expert Rubrics for RLVR

Model ReleasesDGX agent

arXiv:2606.09118v1 Announce Type: new Abstract: As LLM capabilities advance rapidly, the evaluation methods used to assess them increasingly lag behind. Traditional benchmarks relied on programmatic v

CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation

SafetyDGX agent

arXiv:2601.15408v2 Announce Type: replace-cross Abstract: Medical vision-language models can automate the generation of radiology reports but struggle with accurate visual grounding and factual consis

Deep Tree Tensor Networks

Model ReleasesDGX agent

arXiv:2502.09928v2 Announce Type: replace-cross Abstract: Originating in quantum physics, tensor networks (TNs) have been widely adopted as exponential machines and parametric decomposers for recognit

Disjoint Generation of Synthetic Data

ResearchDGX agent

arXiv:2507.19700v2 Announce Type: replace Abstract: We propose a new framework for generating tabular synthetic datasets via disjoint generative models. In this paradigm, a dataset is partitioned into

Frequency-Domain Latent Attention Gating for Cross-Domain Token Aggregation

Model ReleasesDGX agent

arXiv:2606.08191v1 Announce Type: cross Abstract: Token aggregation is a common bottleneck in models that map token representations to sample-level predictions, yet most pooling methods operate only i

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts

Model ReleasesDGX agent

arXiv:2606.09105v1 Announce Type: new Abstract: Generating novel, feasible, and high-quality research ideas is an important yet challenging task in scientific discovery.Recent Large Language Model (LL

HDSL: A Hierarchical Domain-Specific Language for Structured 3D Indoor Scene Generation and Localized Editing with LLM Agents

Model ReleasesDGX agent

arXiv:2606.09738v1 Announce Type: new Abstract: Text-driven indoor scene generation and editing require an intermediate representation that language models can both produce and revise. Existing LLM-ba

iOSWorld: A Benchmark for Personally Intelligent Phone Agents

Model ReleasesDGX agent

arXiv:2606.09764v1 Announce Type: new Abstract: A useful phone agent needs to be personally intelligent. It should reason over a user's identity, history, and preferences as they exist on the device,

Kunlun: Establishing Scaling Laws for Massive-Scale Recommendation Systems through Unified Architecture Design

HardwareDGX agent

arXiv:2602.10016v3 Announce Type: replace-cross Abstract: Deriving predictable scaling laws that govern the relationship between model performance and computational investment is crucial for designing

Language as a Sensor: Calibrated Spatial Belief Estimation in 3D Scenes from Natural Language

Model ReleasesDGX agent

arXiv:2606.08666v1 Announce Type: new Abstract: Robots deployed in human-centric environments routinely receive natural-language descriptions of spatial information ('I left my backpack on the table')

MedVision: Benchmarking Quantitative Medical Image Analysis

Model ReleasesDGX agent

arXiv:2511.18676v2 Announce Type: replace-cross Abstract: Current vision-language models (VLMs) in medicine are primarily designed for categorical question answering (e.g., 'Is this normal or abnormal

Microsoft AI head calls out Anthropic for acting like Claude is conscious

Model ReleasesDGX agent

Microsoft AI CEO Mustafa Suleyman says it's 'really, really dangerous' for Anthropic to speculate about Claude's consciousness inside its 'constitution,' or the instructions that tell the model how to

More Yap Less Meaning: Uncovering Self-Improvement Behavior in SLMs

ResearchDGX agent

arXiv:2606.08471v1 Announce Type: cross Abstract: Recently, language models have made rapid progress across various domains and applications. However, their capability for self-improvement, i.e., whet

Next-Token Prediction Learns Generalisable Representations of Sleep Physiology

ApplicationsDGX agent

arXiv:2606.09605v1 Announce Type: new Abstract: Foundation models offer a promising route to compress multi-modal physiological signals into compact representations of human health, with broad applica

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

Model ReleasesDGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

OmniTryOn: Video Try-On Anything at Once!

Model ReleasesDGX agent

arXiv:2606.08514v1 Announce Type: new Abstract: Although video virtual try-on (VVT) has achieved significant progress, existing methods still exhibit two fundamental limitations: first, they are restr

One Stone, Three Birds: Self-adaptive Optimal Transport for Multi-VLM Selection, Adaptation, and Ensembling

ResearchDGX agent

arXiv:2606.08126v1 Announce Type: new Abstract: Vision-language models (VLMs) enable visual recognition from semantic class descriptions, which makes them attractive when target annotations are scarce

Online Learning with Recency: Algorithms for Sliding-window Streaming Multi-armed Bandits

Model ReleasesDGX agent

arXiv:2606.08977v1 Announce Type: new Abstract: Motivated by the recency effect in online learning, we study algorithms for single-pass *sliding-window streaming multi-armed bandits (MABs)* in this pa

PereStruct: Multimodal Semantic Assembly for Robust Historical Document Parsing

Model ReleasesDGX agent

arXiv:2606.07661v1 Announce Type: new Abstract: Parsing historical documents with complex, non-standard layouts remains a fundamental bottleneck in large-scale archival digitization. Unlike modern typ

POET-X: Memory-efficient LLM Training by Scaling Orthogonal Transformation

Model ReleasesDGX agent

arXiv:2603.05500v2 Announce Type: replace-cross Abstract: Efficient and stable training of large language models (LLMs) remains a core challenge in modern machine learning systems. To address this cha

PriFT: Prior-Support Guided Supervised Fine-Tuning

SafetyDGX agent

arXiv:2606.09396v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is an efficient approach for downstream task adaptation and often serves as the initialization stage for reinforcement le

Programmable Silicon Retina on Pixel Processor Array

Model ReleasesDGX agent

arXiv:2606.08370v1 Announce Type: cross Abstract: Standard dynamic vision sensors approximate retinal processing by detecting temporal contrast changes, offering high speed and high dynamic range. In

Quantum feature-map learning with reduced resource overhead

Model ReleasesDGX agent

arXiv:2510.03389v2 Announce Type: replace-cross Abstract: Current quantum computers require algorithms that use limited resources economically. In quantum machine learning, success hinges on quantum f

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

Model ReleasesDGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?

Model ReleasesDGX agent

arXiv:2606.08063v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in visual understanding, yet their performance degrades significantly un

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

Model ReleasesDGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

SafeRun: Enabling Determinism in LLM Planning for Running

Model ReleasesDGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

Model ReleasesDGX agent

arXiv:2606.08206v1 Announce Type: new Abstract: We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines

Shift-Dependent Asymmetry: Orthogonal Inverse Low-Rank Adaptation for Federated Medical Segmentation

Model ReleasesDGX agent

arXiv:2606.08687v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of segmentation foundation models for medical imaging. However, most federated LoRA m

SOMA: From Surface Observations to Muscle Anatomy

ResearchDGX agent

arXiv:2606.09246v1 Announce Type: new Abstract: With the growing demand for realistic virtual humans, parametric body models have become a cornerstone of modern medicine, sports, and entertainment app

SSAFE: Simple and Strong AI-Generated Image Detection via Frozen Vision Encoders

Model ReleasesDGX agent

arXiv:2606.08634v1 Announce Type: new Abstract: The rapid advancement of generative models has blurred the boundary between synthetic and real imagery, creating an urgent need for reliable deepfake de

Stabilizing On-Policy Distillation for MLLM Reasoning with Global Normalization

Model ReleasesDGX agent

arXiv:2606.09091v1 Announce Type: cross Abstract: On-policy distillation (OPD) has recently emerged as an important post-training paradigm. By using a stronger teacher model to provide dense, fine-gra

Structured Neuron Pruning in Deep Neural Networks Using Multi-Armed Bandits

Model ReleasesDGX agent

arXiv:2606.07615v1 Announce Type: cross Abstract: Deep neural networks often contain redundant hidden units. Removing individual weights can reduce parameter count, but unstructured sparsity is not al

Summarization is Not Dead Yet

SafetyDGX agent

arXiv:2606.08000v1 Announce Type: cross Abstract: The progress of large language models (LLMs) has fueled claims that model-generated summaries rival or even surpass human-written references, raising

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

Model ReleasesDGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

Token Sample Complexity of Attention

Model ReleasesDGX agent

arXiv:2512.10656v3 Announce Type: replace Abstract: As context windows in large language models continue to expand, it is essential to characterize how attention behaves at extreme sequence lengths. W

Training-Free Generalized Few-Shot Segmentation through Open-Vocabulary Semantic Arbitration

Model ReleasesDGX agent

arXiv:2606.09474v1 Announce Type: new Abstract: Generalized Few-Shot Semantic Segmentation (GFSS) has traditionally been approached as a representation-learning problem, requiring task-specific adapta

TUDSR: Twice Upsampling-Diffusion for Higher Super-Resolution

HardwareDGX agent

arXiv:2606.09608v1 Announce Type: new Abstract: Diffusion-based generative models have achieved remarkable success in real-world image super-resolution (SR). With tiled diffusion techniques, these mod

Weak-Driven Learning: How Weak Agents make Strong Agents Stronger

TutorialsDGX agent

arXiv:2602.08222v2 Announce Type: replace Abstract: As post-training optimization becomes central to improving large language models, we observe a persistent saturation bottleneck: once models grow hi

What Codex unlocks for Notion

Model ReleasesDGX agent

OpenAI's Codex model enables Notion to add AI-powered capabilities to its workspace platform, allowing users to automate tasks and generate content through natural language commands. This integration

8 Jun 2026

AdaGRPO: A Capability-Aware Adaptive Enhancement for Flow-based GRPO

SafetyDGX agent

arXiv:2606.06828v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has demonstrated remarkable success in aligning text-to-image (T2I) flow models with human preferences. Howeve

Aumann-SHAP: The Geometry of Counterfactual Interaction Explanations in Machine Learning

Model ReleasesDGX agent

arXiv:2603.14014v2 Announce Type: replace Abstract: We introduce Aumann-SHAP, an interaction-aware framework that decomposes counterfactual transitions by restricting the model to a local hypercube co

Building super fast experiences with Gemma just got easier. Gemma 4 MTP is now officially merged into llama.cpp. Developers can now pair MTP…

Model ReleasesDGX agent

Gemma 4 MTP (Multi-Token Prediction) has been officially integrated into llama.cpp, enabling developers to build faster AI experiences by using the model with this inference framework. This merge allo

CoMetaPNS: Continually Meta-learning Personalized Neural Surrogates for Cardiac Electrophysiology Simulations

ResearchDGX agent

arXiv:2606.07488v1 Announce Type: new Abstract: Personalized virtual heart simulations face challenges in model personalization and computational cost. While neural surrogates offer state-of-the-art s

Compute-Optimal Network Design for Echocardiography Myocardial Segmentation and Perfusion Quantification using Neural Scaling Laws

Model ReleasesDGX agent

arXiv:2606.06725v1 Announce Type: cross Abstract: Myocardial perfusion quantification using contrast-enhanced ultrasound offers a bedside non-ionizing alternative to nuclear imaging modalities. Howeve

CoQuIR: A Comprehensive Benchmark for Code Quality-Aware Information Retrieval

Model ReleasesDGX agent

arXiv:2506.11066v3 Announce Type: replace-cross Abstract: Code retrieval is essential in modern software development, as it boosts code reuse and accelerates debugging. However, current benchmarks pri

DaX: Learning General Pathology Representations Across Scales

Model ReleasesDGX agent

arXiv:2606.06983v1 Announce Type: cross Abstract: Computational pathology requires visual representations that transfer across diverse clinical endpoints and remain robust to variation in magnificatio

Declarative Skills for AI Agents in Knowledge-Grounded Tool-Use Workflows

Model ReleasesDGX agent

arXiv:2606.06923v1 Announce Type: new Abstract: We study orchestration mechanisms for tool-using AI agents in realistic customer-service workflows over an unstructured knowledge base. We argue that de

From Vision to Text: A Compact Multimodal Approach for Robust, Cross-Domain Presentation Attack Detection on ID Cards

Model ReleasesDGX agent

arXiv:2606.06966v1 Announce Type: new Abstract: Cross-domain shifts challenge Presentation Attack Detection (PAD) on ID Cards, given the restricted data available due to privacy concerns. This work pr

Gemma 4 Chat Template now has preserve thinking

Model ReleasesDGX agent

Google added an empty thinking token to the Gemma 4 chat template, which stabilizes model output by suppressing 'ghost' thought channels that may appear even when thinking is deactivated. This update

Learning Perspectivist Social Meaning via Demographic-Conditioned Fusion Embeddings

Model ReleasesDGX agent

arXiv:2606.07123v1 Announce Type: new Abstract: Social meaning in language is inherently perspectival, varying across annotator backgrounds, demographics, and ideological positions. However, most NLP

← Previous
1…397398399400401…1053
Next →