AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,569 results
2 Jul 2026

Radial Interaction Tomography: Recognizing Non-Transitive Evolutionary Games from One Range-Expansion Image

Model ReleasesDGX agent

arXiv:2607.00378v1 Announce Type: new Abstract: Colored sectors in a microbial range expansion encode more than lineage survival counts. We formulate a computer-vision inverse problem: from one endpoi

Rampart, our PII removal model, has cracked the first screen of the top trending models across any category on Huggingface, on the same tier…

Model ReleasesDGX agent

Rampart, our PII removal model, has cracked the first screen of the top trending models across any category on Huggingface, on the same tier as GLM 5.2 / Deepseek! If building systems at fast pace at

RC-GeoCP: Geometric Consensus for Radar-Camera Collaborative Perception


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2603.00654v3 Announce Type: replace Abstract: Collaborative perception (CP) improves scene understanding through multi-agent information sharing, yet LiDAR-centric systems remain costly and vuln

Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models

Model ReleasesDGX agent

arXiv:2607.00852v1 Announce Type: cross Abstract: This work studies the hidden-state inversion problem: recovering the original input token sequence of a decoder-only language model from its last-laye

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

Model ReleasesDGX agent

arXiv:2607.00310v1 Announce Type: cross Abstract: Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic vid

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

Model ReleasesDGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

Model ReleasesDGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

Model ReleasesDGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

Robust 3D Alignment of Generative Reconstructions via Partial Monocular Observations

Model ReleasesDGX agent

arXiv:2607.00498v1 Announce Type: new Abstract: Aligning generative 3D reconstructions with partial monocular observations is a critical but under-explored challenge in computer vision. This task is i

SAOT: Self-Supervised Continual Graph Learning with Structure-Aware Optimal Transport

Model ReleasesDGX agent

arXiv:2607.00377v1 Announce Type: new Abstract: Self-supervised Continual Graph Learning (CGL) aims to successively learn from a graph sequence with different tasks without label supervision - a parad

Seahorse: A Unified Benchmarking Framework for Spatiotemporal Event Modeling

Model ReleasesDGX agent

arXiv:2607.01022v1 Announce Type: new Abstract: Spatiotemporal point processes (STPPs) model event data in continuous time and space, with applications in mobility, epidemiology, and public safety. Re

SegFly: A Dataset and 2D-3D-2D Paradigm for Aerial RGB-Thermal Semantic Segmentation at Scale

Model ReleasesDGX agent

arXiv:2603.17920v2 Announce Type: replace Abstract: Semantic segmentation for uncrewed aerial vehicles (UAVs) is fundamental for aerial scene understanding, yet existing RGB and RGB-T datasets remain

Sheet Music Benchmark: Standardized Optical Music Recognition Evaluation

Model ReleasesDGX agent

arXiv:2506.10488v3 Announce Type: replace Abstract: In this work, we introduce the Sheet Music Benchmark (SMB), a dataset of six hundred and eighty-five pages specifically designed to benchmark Optica

Skills Are Not Islands: Measuring Dependency and Risk in Agent Skill Supply Chains

Model ReleasesDGX agent

arXiv:2607.01136v1 Announce Type: cross Abstract: Agent skills package reusable operational knowledge for Large Language Model (LLM) agents, yet as they grow in scope, they become dependency-bearing a

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to …

Model ReleasesDGX agent

so proud to be working with a bunch of people who are absolutely crushing it! check out all these new things if you haven’t had a chance to yet. huge week for us here @LangChain!! big week at langchai

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models

Model ReleasesDGX agent

arXiv:2603.16859v2 Announce Type: replace Abstract: Omni-modal large language models (OLMs) redefine human-machine interaction by natively integrating audio, vision, and text. However, existing OLM be

Soft Mixture-of-Recursions: Going Deeper with Recursive Vision Transformers

Model ReleasesDGX agent

arXiv:2607.00774v1 Announce Type: new Abstract: Recent recursive Transformer studies have primarily reused shared parameters across computation steps to construct compact, parameter-efficient models.

SoftBank and its telecom unit launch SB Neo to offer AI chips and cloud services to big companies, aiming to provide 10GW of capacity in the US by 2030 (Min-Jeong Lee/Bloomberg)

Model ReleasesDGX agent

Min-Jeong Lee / Bloomberg: SoftBank and its telecom unit launch SB Neo to offer AI chips and cloud services to big companies, aiming to provide 10GW of capacity in the US by 2030 — SoftBank Group Corp

Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an 'order of magnitude more compute than Avocado' (Business Insider)

Model ReleasesDGX agent

Business Insider: Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an “order of magnitude more compute than Avocado” — Meta is making sign

Spectral and Trajectory Regularization for Diffusion Transformer Super-Resolution

Model ReleasesDGX agent

arXiv:2603.06275v2 Announce Type: replace Abstract: Diffusion transformer (DiT) architectures show great potential for real-world image super-resolution (Real-ISR). However, their computationally expe

SpiralFovea: Input-Adaptive Foveated Tokenization as a Third Lever of Resource-Adaptive Inference

Model ReleasesDGX agent

arXiv:2607.00780v1 Announce Type: new Abstract: Most adaptive-inference techniques for foundation models change what the model does - early exit, MoE routing, KV-cache compression, dynamic attention s

Steal the Patch Size: Adversarially Manipulate Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.00174v1 Announce Type: new Abstract: We present a black-box model-stealing attack that recovers private vision-tokenizer configurations of deployed vision-language models (VLMs), including

StochasT: Learning with Stochastic Turn Depth for Visual Instruction Tuning

Model ReleasesDGX agent

arXiv:2607.00465v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely extensively on Visual Instruction Tuning (VIT) to elicit their multimodal reasoning capabilities. However, w

Svarna: An Open Corpus Workbench for Modern Greek

Model ReleasesDGX agent

arXiv:2607.00970v1 Announce Type: new Abstract: This paper introduces Svarna, a free, open-source, web-based corpus workbench for modern Greek. Svarna integrates five databases covering various regist

SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests

Model ReleasesDGX agent

arXiv:2607.00990v1 Announce Type: cross Abstract: Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue re

Tail-Shape Estimation in LLM Evaluation Is Fragile: A Protocol for Diagnosing False Positives

Model ReleasesDGX agent

arXiv:2606.16511v2 Announce Type: replace Abstract: Recent work motivates moving large language model (LLM) evaluation from mean-based to tail-aware metrics, including conditional value-at-risk and ta

TallyTrain: Communication-Efficient Federated Distillation

Model ReleasesDGX agent

arXiv:2607.00173v1 Announce Type: new Abstract: Federated learning is bandwidth-bound on two orthogonal axes: model size, which limits how often parameter-averaging methods can afford to merge, and cl

TANDEM: Temporal Attention-guided Neural Differential Equations for Missingness in Time Series Classification

Model ReleasesDGX agent

arXiv:2508.17519v3 Announce Type: replace-cross Abstract: Handling missing data in time series classification remains a significant challenge in various domains. Traditional methods often rely on impu

TCMA: Text-Conditioned Multi-granularity Alignment for Drone Cross-Modal Text-Video Retrieval

Model ReleasesDGX agent

arXiv:2510.10180v2 Announce Type: replace Abstract: Unmanned aerial vehicles (UAVs) have become powerful platforms for real-time, high-resolution data collection, producing massive volumes of aerial v

TerraBench: Can Agents Reason Over Heterogeneous Earth-System Data?

Model ReleasesDGX agent

arXiv:2606.13148v2 Announce Type: replace Abstract: Climate and environmental decision-making increasingly requires reasoning across heterogeneous inputs, including gridded physical data, satellite im

Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds

Model ReleasesDGX agent

arXiv:2607.00276v1 Announce Type: cross Abstract: Current large-language-model (LLM) physics benchmarks are usually scored by answer accuracy, which cannot distinguish genuine reasoning from recall of

The Download: a startup has a solution for AI’s groupthink problem

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. LLMs are stuck in a groupthink groove. This startup is trying

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology

Model ReleasesDGX agent

arXiv:2607.01033v1 Announce Type: new Abstract: Model organisms (MOs) - language models trained to exhibit undesired or unnatural behaviours - are frequently used as testbeds for evaluating white-box

The talk about Mythos and cybersecurity was not, in fact, hype. (As anyone using Fable to do autonomous work has probably recognized)

Model ReleasesDGX agent

The talk about Mythos and cybersecurity was not, in fact, hype. (As anyone using Fable to do autonomous work has probably recognized) AI appears to be finding software vulnerabilities at scale. In Jun

Timesynth: A Temporal Fidelity Framework for Health Signal Digital Twins

Model ReleasesDGX agent

arXiv:2607.00431v1 Announce Type: new Abstract: Forecasting models for health-signal digital twins must preserve the oscillatory, frequency, phase, and state-transition dynamics of physiological signa

Toward Cybersecurity-Expert Small Language Models

Model ReleasesDGX agent

arXiv:2510.14113v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming everyday applications, yet deployment in cybersecurity lags due to a lack of high-quality, domai

Towards High-Resolution Visual Perception via Hierarchical Entity Exploration

Model ReleasesDGX agent

arXiv:2607.00816v1 Announce Type: new Abstract: High-resolution (HR) image perception remains a key challenge in multimodal large language models (MLLMs), as fine-grained details are often lost when t

Towards Metric-Agnostic Trajectory Forecasting

Model ReleasesDGX agent

arXiv:2607.01133v1 Announce Type: new Abstract: Accurate trajectory forecasting of surrounding traffic participants is a core capability for autonomous driving, enabling vehicles to anticipate behavio

TRIE: An Evaluation Framework for Stochastic PDE Surrogates

Model ReleasesDGX agent

arXiv:2607.00196v1 Announce Type: new Abstract: Many scientific systems exhibit uncertainty from stochastic forcing, unresolved degrees of freedom, or imperfect observations, making reliable surrogate

UltraFlux: Data-Model Co-Design for High-quality Native 4K Text-to-Image Generation across Diverse Aspect Ratios

Model ReleasesDGX agent

arXiv:2511.18050v1 Announce Type: cross Abstract: Diffusion transformers have recently delivered strong text-to-image generation around 1K resolution, but we show that extending them to native 4K acro

Understanding How Humans Inject Knowledge into Machine Learning Workflows through Visual Analytics

Model ReleasesDGX agent

arXiv:2607.00969v1 Announce Type: cross Abstract: Visual analytics (VA) plays an increasingly important role in supporting machine learning (ML) workflows. In the field of visualization, such approach

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.04453v4 Announce Type: replace Abstract: World models have become central to autonomous driving, where accurate scene understanding and future prediction are crucial for safe control. Recen

Using DSPy to evaluate and improve Datasette Agent's SQL system prompts

Model ReleasesDGX agent

Research: Using DSPy to evaluate and improve Datasette Agent's SQL system prompts One of this morning's AIE keynotes covered dspy, which reminded me I've been meaning to see if it could help me improv

Validating Causal Abstraction Metrics on Simulated Complex Systems

Model ReleasesDGX agent

arXiv:2607.00267v1 Announce Type: cross Abstract: A central goal of science is to produce valid explanations of complex systems: high-level causal accounts that faithfully reflect the behavior of lowe

Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

Model ReleasesDGX agent

arXiv:2503.13445v3 Announce Type: replace-cross Abstract: When asked to explain their decisions, LLMs can often give explanations which sound plausible to humans. But are these explanations faithful,

VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2511.17731v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has proven remarkably effective for eliciting complex reasoning in large language models (LLMs). Yet, its potential

VolumeDP: Modeling Volumetric Representation for Manipulation Policy Learning

Model ReleasesDGX agent

arXiv:2603.17720v2 Announce Type: replace Abstract: Imitation learning is a prominent paradigm for robotic manipulation. However, existing visual imitation methods map 2D image observations directly t

Wake up for Touch! Mask-isolated Tactile Alignment Learning in MLLMs

Model ReleasesDGX agent

arXiv:2607.00302v1 Announce Type: new Abstract: Touch supplies the physical grounding needed to perceive intrinsic material properties, such as friction and compliance, that vision alone often cannot

We are hiring our founding team in Korea 🇰🇷 Join us! P.S. Mistral will be at @icmlconf (July 6–11). Come meet the team!

Model ReleasesDGX agent

Mistral AI is recruiting for its founding team in Korea and will have representatives attending ICML conference from July 6-11, 2024, where interested candidates can meet the team in person.

What's Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.00283v1 Announce Type: cross Abstract: Autonomous vehicles must safely navigate complex environments where planning-critical agents may be hidden from view. Current approaches often treat a

Why Advanced Encoders Lag on Sparse Retrieval? The Answer and an Approach to Bridging Vocabulary Gaps

Model ReleasesDGX agent

arXiv:2607.00004v1 Announce Type: cross Abstract: While advanced foundation models like ModernBERT significantly outperform older architectures in dense retrieval, they surprisingly lag behind the agi

Wordle 1,839 4/6 ⬛⬛🟨⬛🟨 ⬛⬛🟨⬛⬛ ⬛🟨⬛🟨🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

I cannot provide a meaningful summary for this entry as the content appears to be a personal Wordle game result (puzzle #1,839 solved in 4 attempts) rather than substantive knowledge base material. Th

WorkBench Revisited: Workplace Agents Two Years On

Model ReleasesDGX agent

arXiv:2606.13715v2 Announce Type: replace Abstract: The best agent on WorkBench in March 2024, GPT-4, completed just 43% of tasks. We revisit the benchmark in June 2026 and find that the best agent to

XSkill: Continual Learning from Experience and Skills in Multimodal Agents

Model ReleasesDGX agent

arXiv:2603.12056v3 Announce Type: replace Abstract: Multimodal agents can now tackle complex reasoning tasks with diverse tools, yet they still suffer from inefficient tool use and inflexible orchestr

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese

Model ReleasesDGX agent

arXiv:2607.00664v1 Announce Type: new Abstract: We propose YOMI-Bench, a benchmark for evaluating kanji reading and phonological understanding of large language models (LLMs) for Japanese. In Japanese

You really need your own benchmarks. If you are translating hieroglyphics, use Gemini 3.5 Flash. If you are running a vending machine use Op…

Model ReleasesDGX agent

You really need your own benchmarks. If you are translating hieroglyphics, use Gemini 3.5 Flash. If you are running a vending machine use Opus 4.8. (This is one reason why I am skeptical of just swapp

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in d…

Model ReleasesDGX agent

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in different formats. The second your team uses more than one (t

Z.ai launches ZCode, an 'Agentic Development Environment' optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month (Michael Nuñez/VentureBeat)

Model ReleasesDGX agent

Michael Nuñez / VentureBeat: Z.ai launches ZCode, an “Agentic Development Environment” optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month — The move marks th

ZO-Act: Efficient Zeroth-Order Fine-Tuning via One-Shot Activation-Informed Low-Rank Subspaces

Model ReleasesDGX agent

arXiv:2607.01125v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization enables fine-tuning large language models when backpropagation is unavailable or memory-prohibitive, but existing methods

1 Jul 2026

A Large-Language-Model Supported Personalized Driving Framework for Lane Change in Highway Scenarios

Model ReleasesDGX agent

arXiv:2606.31483v1 Announce Type: new Abstract: Personalized driving can improve the user acceptance of automated driving systems. However, existing methods still provide limited support for translati

← Previous
1…111112113114115…377
Next →