AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
2 Jul 2026

Persona Without Substrate: Regime-Dependence and the LLM Individuation Problem

Model ReleasesDGX agent

arXiv:2607.00006v1 Announce Type: cross Abstract: Beckmann & Butlin's (2026) ontological framework for the LLM individuation problem inherits an unargued cross-regime co-reference assumption from the

PHREEQC-MCQ-200: A Diagnostic Benchmark for Tool-Augmented Scientific Simulator Agents

Model ReleasesDGX agent

arXiv:2607.00436v1 Announce Type: new Abstract: Large language model agents are increasingly connected to scientific software, yet it remains unclear when tool access makes scientific computation more

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

PorTEXTO: A European Portuguese Benchmark for Visual Text Extraction

Model ReleasesDGX agent

arXiv:2606.19096v2 Announce Type: replace Abstract: European Portuguese (pt-PT) is largely absent from OCR benchmarks, which skew toward high-resource languages. The few benchmarks that cover pt-PT fo

Post-Training Pruning for Diffusion Transformers

Model ReleasesDGX agent

arXiv:2607.00927v1 Announce Type: cross Abstract: Diffusion Transformers (DiTs) have demonstrated impressive performance in image generation but suffer from substantial computational overhead and reso

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

Model ReleasesDGX agent

arXiv:2607.00049v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt the

QuaMoE-DRF: Proactive Beam and Rate Adaptation via Multimodal Dynamic Radio Map Forecasting in ISAC Networks

Model ReleasesDGX agent

arXiv:2607.00974v1 Announce Type: cross Abstract: Static radio maps provide location-dependent propagation priors, but they cannot capture short-term blockage caused by moving objects. Direct sensing-

Quantifying the Affective Gap: A Zero-Shot Evaluation of LLMs on Fine-Grained Emotion Taxonomies

Model ReleasesDGX agent

arXiv:2607.00968v1 Announce Type: new Abstract: Emotion recognition in natural language is a foundational challenge in affective computing, with critical implications for human-computer interaction, m

Quantum vs. Classical Machine Learning: A Unified Empirical Comparison

Model ReleasesDGX agent

arXiv:2607.01197v1 Announce Type: new Abstract: Quantum computing has emerged as a promising computational paradigm for machine learning (ML), with the potential to offer computational advantages over

Radial Interaction Tomography: Recognizing Non-Transitive Evolutionary Games from One Range-Expansion Image

Model ReleasesDGX agent

arXiv:2607.00378v1 Announce Type: new Abstract: Colored sectors in a microbial range expansion encode more than lineage survival counts. We formulate a computer-vision inverse problem: from one endpoi

Rampart, our PII removal model, has cracked the first screen of the top trending models across any category on Huggingface, on the same tier…

Model ReleasesDGX agent

Rampart, our PII removal model, has cracked the first screen of the top trending models across any category on Huggingface, on the same tier as GLM 5.2 / Deepseek! If building systems at fast pace at

RC-GeoCP: Geometric Consensus for Radar-Camera Collaborative Perception

Model ReleasesDGX agent

arXiv:2603.00654v3 Announce Type: replace Abstract: Collaborative perception (CP) improves scene understanding through multi-agent information sharing, yet LiDAR-centric systems remain costly and vuln

Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models

Model ReleasesDGX agent

arXiv:2607.00852v1 Announce Type: cross Abstract: This work studies the hidden-state inversion problem: recovering the original input token sequence of a decoder-only language model from its last-laye

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

Model ReleasesDGX agent

arXiv:2607.00310v1 Announce Type: cross Abstract: Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic vid

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

Model ReleasesDGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

Model ReleasesDGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

Model ReleasesDGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

Robust 3D Alignment of Generative Reconstructions via Partial Monocular Observations

Model ReleasesDGX agent

arXiv:2607.00498v1 Announce Type: new Abstract: Aligning generative 3D reconstructions with partial monocular observations is a critical but under-explored challenge in computer vision. This task is i

SAOT: Self-Supervised Continual Graph Learning with Structure-Aware Optimal Transport

Model ReleasesDGX agent

arXiv:2607.00377v1 Announce Type: new Abstract: Self-supervised Continual Graph Learning (CGL) aims to successively learn from a graph sequence with different tasks without label supervision - a parad

Seahorse: A Unified Benchmarking Framework for Spatiotemporal Event Modeling

Model ReleasesDGX agent

arXiv:2607.01022v1 Announce Type: new Abstract: Spatiotemporal point processes (STPPs) model event data in continuous time and space, with applications in mobility, epidemiology, and public safety. Re

SegFly: A Dataset and 2D-3D-2D Paradigm for Aerial RGB-Thermal Semantic Segmentation at Scale

Model ReleasesDGX agent

arXiv:2603.17920v2 Announce Type: replace Abstract: Semantic segmentation for uncrewed aerial vehicles (UAVs) is fundamental for aerial scene understanding, yet existing RGB and RGB-T datasets remain

Sheet Music Benchmark: Standardized Optical Music Recognition Evaluation

Model ReleasesDGX agent

arXiv:2506.10488v3 Announce Type: replace Abstract: In this work, we introduce the Sheet Music Benchmark (SMB), a dataset of six hundred and eighty-five pages specifically designed to benchmark Optica

Skills Are Not Islands: Measuring Dependency and Risk in Agent Skill Supply Chains

Model ReleasesDGX agent

arXiv:2607.01136v1 Announce Type: cross Abstract: Agent skills package reusable operational knowledge for Large Language Model (LLM) agents, yet as they grow in scope, they become dependency-bearing a

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to …

Model ReleasesDGX agent

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to yet. huge week for us here @LangChain!! big week at langchai

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models

Model ReleasesDGX agent

arXiv:2603.16859v2 Announce Type: replace Abstract: Omni-modal large language models (OLMs) redefine human-machine interaction by natively integrating audio, vision, and text. However, existing OLM be

Soft Mixture-of-Recursions: Going Deeper with Recursive Vision Transformers

Model ReleasesDGX agent

arXiv:2607.00774v1 Announce Type: new Abstract: Recent recursive Transformer studies have primarily reused shared parameters across computation steps to construct compact, parameter-efficient models.

SoftBank and its telecom unit launch SB Neo to offer AI chips and cloud services to big companies, aiming to provide 10GW of capacity in the US by 2030 (Min-Jeong Lee/Bloomberg)

Model ReleasesDGX agent

Min-Jeong Lee / Bloomberg: SoftBank and its telecom unit launch SB Neo to offer AI chips and cloud services to big companies, aiming to provide 10GW of capacity in the US by 2030 — SoftBank Group Corp

Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an 'order of magnitude more compute than Avocado' (Business Insider)

Model ReleasesDGX agent

Business Insider: Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an “order of magnitude more compute than Avocado” — Meta is making sign

Spectral and Trajectory Regularization for Diffusion Transformer Super-Resolution

Model ReleasesDGX agent

arXiv:2603.06275v2 Announce Type: replace Abstract: Diffusion transformer (DiT) architectures show great potential for real-world image super-resolution (Real-ISR). However, their computationally expe

SpiralFovea: Input-Adaptive Foveated Tokenization as a Third Lever of Resource-Adaptive Inference

Model ReleasesDGX agent

arXiv:2607.00780v1 Announce Type: new Abstract: Most adaptive-inference techniques for foundation models change what the model does - early exit, MoE routing, KV-cache compression, dynamic attention s

Steal the Patch Size: Adversarially Manipulate Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.00174v1 Announce Type: new Abstract: We present a black-box model-stealing attack that recovers private vision-tokenizer configurations of deployed vision-language models (VLMs), including

StochasT: Learning with Stochastic Turn Depth for Visual Instruction Tuning

Model ReleasesDGX agent

arXiv:2607.00465v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely extensively on Visual Instruction Tuning (VIT) to elicit their multimodal reasoning capabilities. However, w

Svarna: An Open Corpus Workbench for Modern Greek

Model ReleasesDGX agent

arXiv:2607.00970v1 Announce Type: new Abstract: This paper introduces Svarna, a free, open-source, web-based corpus workbench for modern Greek. Svarna integrates five databases covering various regist

SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests

Model ReleasesDGX agent

arXiv:2607.00990v1 Announce Type: cross Abstract: Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue re

Tail-Shape Estimation in LLM Evaluation Is Fragile: A Protocol for Diagnosing False Positives

Model ReleasesDGX agent

arXiv:2606.16511v2 Announce Type: replace Abstract: Recent work motivates moving large language model (LLM) evaluation from mean-based to tail-aware metrics, including conditional value-at-risk and ta

TallyTrain: Communication-Efficient Federated Distillation

Model ReleasesDGX agent

arXiv:2607.00173v1 Announce Type: new Abstract: Federated learning is bandwidth-bound on two orthogonal axes: model size, which limits how often parameter-averaging methods can afford to merge, and cl

TANDEM: Temporal Attention-guided Neural Differential Equations for Missingness in Time Series Classification

Model ReleasesDGX agent

arXiv:2508.17519v3 Announce Type: replace-cross Abstract: Handling missing data in time series classification remains a significant challenge in various domains. Traditional methods often rely on impu

TCMA: Text-Conditioned Multi-granularity Alignment for Drone Cross-Modal Text-Video Retrieval

Model ReleasesDGX agent

arXiv:2510.10180v2 Announce Type: replace Abstract: Unmanned aerial vehicles (UAVs) have become powerful platforms for real-time, high-resolution data collection, producing massive volumes of aerial v

TerraBench: Can Agents Reason Over Heterogeneous Earth-System Data?

Model ReleasesDGX agent

arXiv:2606.13148v2 Announce Type: replace Abstract: Climate and environmental decision-making increasingly requires reasoning across heterogeneous inputs, including gridded physical data, satellite im

Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds

Model ReleasesDGX agent

arXiv:2607.00276v1 Announce Type: cross Abstract: Current large-language-model (LLM) physics benchmarks are usually scored by answer accuracy, which cannot distinguish genuine reasoning from recall of

The Download: a startup has a solution for AI’s groupthink problem

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. LLMs are stuck in a groupthink groove. This startup is trying

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology

Model ReleasesDGX agent

arXiv:2607.01033v1 Announce Type: new Abstract: Model organisms (MOs) - language models trained to exhibit undesired or unnatural behaviours - are frequently used as testbeds for evaluating white-box

The talk about Mythos and cybersecurity was not, in fact, hype. (As anyone using Fable to do autonomous work has probably recognized)

Model ReleasesDGX agent

The talk about Mythos and cybersecurity was not, in fact, hype. (As anyone using Fable to do autonomous work has probably recognized) AI appears to be finding software vulnerabilities at scale. In Jun

Timesynth: A Temporal Fidelity Framework for Health Signal Digital Twins

Model ReleasesDGX agent

arXiv:2607.00431v1 Announce Type: new Abstract: Forecasting models for health-signal digital twins must preserve the oscillatory, frequency, phase, and state-transition dynamics of physiological signa

Toward Cybersecurity-Expert Small Language Models

Model ReleasesDGX agent

arXiv:2510.14113v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming everyday applications, yet deployment in cybersecurity lags due to a lack of high-quality, domai

Towards High-Resolution Visual Perception via Hierarchical Entity Exploration

Model ReleasesDGX agent

arXiv:2607.00816v1 Announce Type: new Abstract: High-resolution (HR) image perception remains a key challenge in multimodal large language models (MLLMs), as fine-grained details are often lost when t

Towards Metric-Agnostic Trajectory Forecasting

Model ReleasesDGX agent

arXiv:2607.01133v1 Announce Type: new Abstract: Accurate trajectory forecasting of surrounding traffic participants is a core capability for autonomous driving, enabling vehicles to anticipate behavio

TRIE: An Evaluation Framework for Stochastic PDE Surrogates

Model ReleasesDGX agent

arXiv:2607.00196v1 Announce Type: new Abstract: Many scientific systems exhibit uncertainty from stochastic forcing, unresolved degrees of freedom, or imperfect observations, making reliable surrogate

UltraFlux: Data-Model Co-Design for High-quality Native 4K Text-to-Image Generation across Diverse Aspect Ratios

Model ReleasesDGX agent

arXiv:2511.18050v1 Announce Type: cross Abstract: Diffusion transformers have recently delivered strong text-to-image generation around 1K resolution, but we show that extending them to native 4K acro

Understanding How Humans Inject Knowledge into Machine Learning Workflows through Visual Analytics

Model ReleasesDGX agent

arXiv:2607.00969v1 Announce Type: cross Abstract: Visual analytics (VA) plays an increasingly important role in supporting machine learning (ML) workflows. In the field of visualization, such approach

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.04453v4 Announce Type: replace Abstract: World models have become central to autonomous driving, where accurate scene understanding and future prediction are crucial for safe control. Recen

Using DSPy to evaluate and improve Datasette Agent's SQL system prompts

Model ReleasesDGX agent

Research: Using DSPy to evaluate and improve Datasette Agent's SQL system prompts One of this morning's AIE keynotes covered dspy, which reminded me I've been meaning to see if it could help me improv

Validating Causal Abstraction Metrics on Simulated Complex Systems

Model ReleasesDGX agent

arXiv:2607.00267v1 Announce Type: cross Abstract: A central goal of science is to produce valid explanations of complex systems: high-level causal accounts that faithfully reflect the behavior of lowe

Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

Model ReleasesDGX agent

arXiv:2503.13445v3 Announce Type: replace-cross Abstract: When asked to explain their decisions, LLMs can often give explanations which sound plausible to humans. But are these explanations faithful,

VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2511.17731v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has proven remarkably effective for eliciting complex reasoning in large language models (LLMs). Yet, its potential

VolumeDP: Modeling Volumetric Representation for Manipulation Policy Learning

Model ReleasesDGX agent

arXiv:2603.17720v2 Announce Type: replace Abstract: Imitation learning is a prominent paradigm for robotic manipulation. However, existing visual imitation methods map 2D image observations directly t

Wake up for Touch! Mask-isolated Tactile Alignment Learning in MLLMs

Model ReleasesDGX agent

arXiv:2607.00302v1 Announce Type: new Abstract: Touch supplies the physical grounding needed to perceive intrinsic material properties, such as friction and compliance, that vision alone often cannot

We are hiring our founding team in Korea 🇰🇷 Join us! P.S. Mistral will be at @icmlconf (July 6–11). Come meet the team!

Model ReleasesDGX agent

Mistral AI is recruiting for its founding team in Korea and will have representatives attending ICML conference from July 6-11, 2024, where interested candidates can meet the team in person.

What's Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.00283v1 Announce Type: cross Abstract: Autonomous vehicles must safely navigate complex environments where planning-critical agents may be hidden from view. Current approaches often treat a

Why Advanced Encoders Lag on Sparse Retrieval? The Answer and an Approach to Bridging Vocabulary Gaps

Model ReleasesDGX agent

arXiv:2607.00004v1 Announce Type: cross Abstract: While advanced foundation models like ModernBERT significantly outperform older architectures in dense retrieval, they surprisingly lag behind the agi

← Previous
1…110111112113114…377
Next →