AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,574 results
30 Jun 2026

The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.28843v1 Announce Type: cross Abstract: Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task. However, prior work has shown th

The Hidden Cost of Structured Generation in LLMs: Draft-Conditioned Constrained Decoding

Model ReleasesDGX agent

arXiv:2603.03305v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to generate executable outputs, JSON objects, and API calls, where a single syntax error ca

The Human Creativity Benchmark

Model ReleasesDGX agent

arXiv:2606.30561v1 Announce Type: new Abstract: Modern AI evaluation frameworks treat evaluator disagreement as noise to be resolved. In creative domains, professional disagreement reflects genuine di


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Interference Gap: Comparing Retrieval Bounds in Human Memory and RAG Systems

Model ReleasesDGX agent

arXiv:2606.28327v1 Announce Type: cross Abstract: How do retrieval bounds compare between human episodic memory and Retrieval-Augmented Generation (RAG) systems under semantic interference? We present

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by h…

Model ReleasesDGX agent

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by how strongly GLM-5.2 is pushing long-horizon performance .. not just

The NTNU System at the S&I Challenge 2025 SLA Open Track

Model ReleasesDGX agent

arXiv:2506.05121v3 Announce Type: replace Abstract: A recent line of research on spoken language assessment (SLA) employs neural models such as BERT and wav2vec 2.0 (W2V) to evaluate speaking proficie

The Verbose Context Problem in Medical Records

Model ReleasesDGX agent

arXiv:2606.29503v1 Announce Type: cross Abstract: The verbose context problem occurs when structured concepts have token-inefficient textual representations. This bottleneck is acute in population hea

this is what you look like with low rise pants

Model ReleasesDGX agent

This post likely showcases visual examples or commentary on the aesthetic appearance and fit of low-rise pants, a fashion trend that was particularly popular in the early 2000s and has experienced per

Thrilled to announce the Wearable AI Workshop at ECCV 2026 🎉 that we're organizing with an awesome group of folks across Meta Reality Labs,…

Model ReleasesDGX agent

Thrilled to announce the Wearable AI Workshop at ECCV 2026 🎉 that we're organizing with an awesome group of folks across Meta Reality Labs, AMI Labs, HKUST, Georgia Tech, UCF, and U. of Edinburgh. If

Thunder-KoNUBench: A Corpus-Aligned Benchmark for Korean Negation Understanding

Model ReleasesDGX agent

arXiv:2601.04693v2 Announce Type: replace Abstract: Although negation is known to challenge large language models (LLMs), benchmarks for evaluating negation understanding-especially in Korean-are scar

Toward an Energy-Optimized Operation of Data Centers Located in Wind Farms Using Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.30316v1 Announce Type: new Abstract: This paper studies Reinforcement Learning as an online controller for curtailment-aware workload shifting in wind-turbine-integrated high-performance co

Toward Secure and Reliable PDDL Formalization of Large Language Models with Planner-in-the-Loop Feedback

Model ReleasesDGX agent

arXiv:2606.29700v1 Announce Type: new Abstract: Planning often requires symbolic specifications that are both executable and verifiable. For large language models deployed in autonomous or decision-su

Towards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.30266v1 Announce Type: cross Abstract: Motion-language agents must possess the bidirectional capability to both understand human movement (motion-to-text, M2T) and generate it from natural

Towards Generalizable and Evidential Nuclear Magnetic Resonance-Based Molecular Structure Elucidation via Large Language Model Agent

Model ReleasesDGX agent

arXiv:2606.29776v1 Announce Type: cross Abstract: Nuclear Magnetic Resonance (NMR) spectroscopy is the gold standard for molecular structure elucidation, yet interpreting complex spectra for unknown m

Towards in-the-wild Egocentric 3D Hand-Object Pose Estimation

Model ReleasesDGX agent

arXiv:2606.30598v1 Announce Type: new Abstract: Estimating accurate 3D hand-object pose from in-the-wild egocentric RGB remains challenging due to severe occlusions and ambiguous contact. Existing lea

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics

Model ReleasesDGX agent

arXiv:2512.13660v3 Announce Type: replace-cross Abstract: Spatial tracing, as a fundamental embodied interaction ability for robots, is inherently challenging as it requires multi-step metric-grounded

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

Model ReleasesDGX agent

arXiv:2606.30560v1 Announce Type: cross Abstract: Coding agents are rapidly becoming a major application of agentic LLMs, but serving them efficiently remains challenging. Progress on this challenge r

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision

Model ReleasesDGX agent

arXiv:2606.30552v1 Announce Type: cross Abstract: Cross-embodiment transfer in vision-language-action (VLA) models remains challenging because low-level state and action spaces differ fundamentally ac

Translating Natural Language to Strategic Temporal Specifications via LLMs

Model ReleasesDGX agent

arXiv:2606.30441v1 Announce Type: cross Abstract: A rigorous formalization of system requirements is a fundamental prerequisite for the verification of Multi-Agent Systems (MAS). However, writing corr

Travel-Oriented Reasoning Large Language Model via Domain-Specific Knowledge Graphs

Model ReleasesDGX agent

arXiv:2606.29254v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate broad reasoning abilities but struggle with accuracy and reliability in specialized domains such as travel, whe

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

Model ReleasesDGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

Model ReleasesDGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

Tumor-aware augmentation with task-guided attention analysis improves rectal cancer segmentation from magnetic resonance images

Model ReleasesDGX agent

arXiv:2605.05522v2 Announce Type: replace-cross Abstract: Although self-supervised pretraining is expected to learn broadly transferable representations, its effectiveness across imaging modalities su

Tutorial on using Gemini live to build a voice agent Uses deepagents as a tool: offload complex work to this subagent, use Gemini live for t…

Model ReleasesDGX agent

Tutorial on using Gemini live to build a voice agent Uses deepagents as a tool: offload complex work to this subagent, use Gemini live for the naturalness/latency Building voice agents can come with t

Two kinds of robustness are not the same: disentangling fault tolerance and low-SNR robustness in multi-domain event detection on real data

Model ReleasesDGX agent

arXiv:2606.29339v1 Announce Type: cross Abstract: Reliable event detection underpins induced-seismicity monitoring for Carbon dioxide Capture and Storage (CCS) and geothermal operations, distributed a

UniCA: Bi-directional Cross-Attention with Positive Similarity Loss for Robust Multi-Modal Retrieval

Model ReleasesDGX agent

arXiv:2606.28350v1 Announce Type: cross Abstract: Multi-modal retrieval has become increasingly critical for handling the growing volume of integrated visual-textual data in real-world applications, b

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

Model ReleasesDGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

Unlocking the Visual Record of Materials Science: A Large-Scale Multimodal Dataset from Scientific Literature

Model ReleasesDGX agent

arXiv:2606.29667v1 Announce Type: cross Abstract: The materials science literature encodes decades of experimental knowledge in figures, yet this visual record remains locked away and inaccessible to

UrbanCDNet: Appearance-Robust and Boundary-Aware Bitemporal Change Detection for Korean Urban Building Monitoring

Model ReleasesDGX agent

arXiv:2606.29781v1 Announce Type: new Abstract: Urban building change detection from bi-temporal aerial imagery is important for redevelopment monitoring, infrastructure management, and unauthorized-c

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

Model ReleasesDGX agent

arXiv:2606.29964v1 Announce Type: new Abstract: Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view:

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

Model ReleasesDGX agent

arXiv:2603.21526v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However,

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

Model ReleasesDGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

We are making a deliberate effort to use GLM 5.2 with OpenCode internally at Jarvislabs. I have spoken to 3 enterprise customers last week, …

Model ReleasesDGX agent

We are making a deliberate effort to use GLM 5.2 with OpenCode internally at Jarvislabs. I have spoken to 3 enterprise customers last week, who are exploring to host multiple open source models and mo

We build differently than other tech firms. No superintelligence. No competition to spend the most money. We're creating empowering, efficie…

Model ReleasesDGX agent

We build differently than other tech firms. No superintelligence. No competition to spend the most money. We're creating empowering, efficient AI to enhance human potential, not replace it. With high

We just made it a lot easier for AI agents to work with your documents. LlamaParse MCP now does more than parse or classify files: it can pu…

Model ReleasesDGX agent

We just made it a lot easier for AI agents to work with your documents. LlamaParse MCP now does more than parse or classify files: it can pull structured data out of contracts, invoices, and reports a

We're excited to be sponsoring and co-organizing IOL-AI 2026, a new open challenge with the International Linguistics Olympiad @IOLing_offic…

Model ReleasesDGX agent

We're excited to be sponsoring and co-organizing IOL-AI 2026, a new open challenge with the International Linguistics Olympiad @IOLing_official.🌎💬 🌐 Open-science 🗓️ One-month competition 🎯 Targeting A

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological …

Model ReleasesDGX agent

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological data, choose the right analysis path, and make judgment call

We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then in…

Model ReleasesDGX agent

We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then instantly animate them with the other—all at a fraction of the

We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring acces…

Model ReleasesDGX agent

We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring access tomorrow, and will share an update soon. We’re grateful to

What an honor to emcee the first day of @aiDotEngineer and introduce the Software Factories Track Thank you @swyx & team, and @KeycardLabs f…

Model ReleasesDGX agent

What an honor to emcee the first day of @aiDotEngineer and introduce the Software Factories Track Thank you @swyx & team, and @KeycardLabs for the support. “A year ago @GeoffreyHuntley released the Ra

What Drives the Inlier-Memorization Effect? A Theory of Outlier Detection via Early Training Dynamics

Model ReleasesDGX agent

arXiv:2606.29791v1 Announce Type: cross Abstract: Outlier detection (OD) aims to identify anomalous instances by learning the underlying structure of normal data (inliers), and is particularly challen

What's new in Claude Sonnet 5

Model ReleasesDGX agent

What's new in Claude Sonnet 5 Claude Sonnet 5 came out this morning. I always head straight for the 'what's new' developer docs because they tend to have more actionable information than the official

When AI Reviews Its Own Code: Recursive Self-Training Collapse in Code LLMs

Model ReleasesDGX agent

arXiv:2606.28438v1 Announce Type: cross Abstract: Recursive self-training can degrade neural generative models when generated data is reused without fresh human data or external quality control. We st

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

Model ReleasesDGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

When Medical Safety Alignment Fails: A Benchmark for Evaluating LLMs on High-Risk Medical Queries

Model ReleasesDGX agent

arXiv:2606.28332v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for medical and health-related questions, yet their safety in high-risk medical scenarios remains p

When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling

Model ReleasesDGX agent

arXiv:2606.28661v1 Announce Type: cross Abstract: People overthink; language models over-sample, and the extra effort can talk both into a worse answer. Reasoning systems answer a hard question by sam

Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents

Model ReleasesDGX agent

arXiv:2606.30383v1 Announce Type: new Abstract: A rapidly growing class of LLM agents is multi-party: the agent acts for a principal (who briefs it, sends follow-ups, and receives results) while also

Why Do We Need Warm-up? A Theoretical Perspective

Model ReleasesDGX agent

arXiv:2510.03164v2 Announce Type: replace Abstract: Learning rate warm-up -- increasing the learning rate at the beginning of training -- has become a ubiquitous heuristic in modern deep learning, yet

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression

Model ReleasesDGX agent

arXiv:2606.29712v1 Announce Type: new Abstract: Large language models achieve high reasoning performance via explicit chain-of-thought and reinforcement learning, but require long output sequences and

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

Model ReleasesDGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

Wordle 1,836 4/6 ⬛⬛⬛🟨🟩 ⬛⬛⬛⬛⬛ ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result (puzzle #1,836) solved in 4 attempts, showing the guess sequence and color-coded feedback (gray for incorrect letters, yellow for correct letters in wrong posi

XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2412.15529v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) synergizes the retrieval of pertinent data with the generative capabilities of Large Language Models (LLM

XYZ-IBD: Benchmarking Robust 6D Object Pose Estimation under Real-World Industrial Complexity

Model ReleasesDGX agent

arXiv:2506.00599v3 Announce Type: replace Abstract: While current 6D pose estimation benchmarks have reached near-saturation on household objects, they often fail to capture the stochastic and optical

You asked, we listened. Claude Desktop on Linux is here! Download link: https://code.claude.com/docs/en/desktop-linux

Model ReleasesDGX agent

You asked, we listened. Claude Desktop on Linux is here! Download link: https://code.claude.com/docs/en/desktop-linux Claude Desktop is now available on Linux (Ubuntu and Debian) in beta. Alongside th

You cannot ban Chinese open source to make US win. The whole point of open source is it's OPEN. You might ban in the US, sure, but how are y…

Model ReleasesDGX agent

You cannot ban Chinese open source to make US win. The whole point of open source is it's OPEN. You might ban in the US, sure, but how are you going to ban it outside US? Instead the real question is,

You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact

Model ReleasesDGX agent

arXiv:2606.28899v1 Announce Type: new Abstract: Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and re

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot.

Model ReleasesDGX agent

You wake up. It’s 2026. The newest open source AI model was trained by federal government employees, Big Balls and Groot. FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr)

Zenith took GPT-5.5 from 5th to 1st on Frontier SWE. @EMostaque and @PeterDiamandis talk it through, and more, on Moonshots. https://youtu.b…

Model ReleasesDGX agent

Zenith took GPT-5.5 from 5th to 1st on Frontier SWE. @EMostaque and @PeterDiamandis talk it through, and more, on Moonshots. https://youtu.be/-H7J_-zr7pA?is=Oeo7gg2ZevKCw2IE Media You don't need Fable

Zero-Gated Language-conditioned Human Motion Prediction

Model ReleasesDGX agent

arXiv:2606.29208v1 Announce Type: new Abstract: Pose histories provide the core kinematic evidence for 3D human motion prediction, but they lack explicit high-level semantic guidance. This paper intro

← Previous
1…123124125126127…377
Next →