AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

Atomistic Language Models Understand and Generate Materials

DGX agent

arXiv:2606.21395v1 Announce Type: new Abstract: Atomistic structure and natural language have long been modeled separately, with language models either calling atomistic models as tools or being fine-

model-releasesarxiv-cs-lg
23 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Attacking the Trusted Imagination: Oracle-Level Integrity Attacks on Imagine-then-Act World Models

DGX agent

arXiv:2606.22966v1 Announce Type: new Abstract: Many recent vision-language-action (VLA) policies adopt an imagine-then-act design. A world-action model (WAM) first imagines a short future as a latent

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

AutoDex: An Automated Real-World System for Dexterous Grasping Data Collection

DGX agent

arXiv:2606.23689v1 Announce Type: cross Abstract: Learning robust dexterous grasping requires real-world data that records the physical outcomes of grasp attempts. Such data is hard to obtain at scale

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Available today: the API, Document AI in Mistral AI Studio, Amazon SageMaker, Microsoft Foundry, coming soon Snowflake Parse Document, or se…

DGX agent

Available today: the API, Document AI in Mistral AI Studio, Amazon SageMaker, Microsoft Foundry, coming soon Snowflake Parse Document, or self-hosted on a single container, so your documents never lea

model-releasesmistral-ai--x
23 Jun 2026
Model Releases

Bayesian Adaptation Gym: A Benchmark for the Bayesian Low-Rank Adaptation of Multi-Modal Language Models

DGX agent

arXiv:2606.22188v1 Announce Type: new Abstract: Large multi-modal language models are increasingly deployed in high-stakes domains, making well-calibrated uncertainty essential. Traditional Bayesian m

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

BELDE: Building a Large-scale Earth-observation Land-cover Dataset for Europe

DGX agent

arXiv:2606.20909v1 Announce Type: new Abstract: Earth observation imagery plays a critical role in environmental monitoring, urban planning, disaster assessment, and climate analysis. While multi-spec

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

BELLS-O: Evaluating the Operational Trade-offs of LLM Supervision Systems

DGX agent

arXiv:2606.20668v1 Announce Type: cross Abstract: LLM supervision systems, namely input/output moderation filters and jailbreak detectors, are the primary safeguard against misuse in deployed AI appli

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Benchmarking Robot Memory Under Interference

DGX agent

arXiv:2606.22338v1 Announce Type: cross Abstract: Robots deployed in realistic settings will accumulate experience across many sessions and tasks over their deployment. The robot's tasks may often req

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Benchmarking Vision-Language Models for Microscopic Plant Image Understanding

DGX agent

arXiv:2606.22497v1 Announce Type: new Abstract: Microscopic imaging provides essential visual evidence for studying plant biology and pathology at the cellular and subcellular levels. However, existin

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation

DGX agent

arXiv:2606.15127v2 Announce Type: replace Abstract: Reasoning models are increasingly used in settings where the final answer is not the only object of review: educational tools may show students inte

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies

DGX agent

arXiv:2606.20599v1 Announce Type: cross Abstract: Tree of Thought (ToT) search has become a promising direction for improving the reasoning capabilities of large language models, but deploying these m

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Beyond 'One Language, One Script': Quantifying Orthographic Bias in Multilingual VLMs with PuMVR

DGX agent

arXiv:2606.20770v1 Announce Type: cross Abstract: Current Vision-Language Models (VLMs) are celebrated for their multilingual capabilities, yet they operate under a flawed assumption: that one languag

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Beyond the LUMIR challenge: The pathway to foundational registration models

DGX agent

arXiv:2505.24160v3 Announce Type: replace-cross Abstract: Medical image challenges have played a transformative role in advancing the field, catalyzing innovation and establishing new performance benc

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation

DGX agent

arXiv:2511.22973v2 Announce Type: replace Abstract: Long video generation is a critical step toward building realistic world models, requiring both high visual fidelity and long-range interaction cons

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Black-Box Continual Learning for Vision-Language Models

DGX agent

arXiv:2606.22999v1 Announce Type: new Abstract: The rapid deployment of Vision-Language Models (VLMs) in dynamic environments necessitates the ability to learn continuously without forgetting. However

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

BranchShine: Compact Raw-Audio-to-IPA Transcription with a RoPE E-Branchformer Encoder

DGX agent

arXiv:2606.22824v1 Announce Type: new Abstract: Speech-to-IPA transcription is useful when the desired output is pronunciation rather than orthographic text, but competitive multilingual systems are o

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Bring open models to where you're already working. FireConnect brings Fireworks' top models directly into Claude Code, Pi, OpenCode, and Cod…

DGX agent

Bring open models to where you're already working. FireConnect brings Fireworks' top models directly into Claude Code, Pi, OpenCode, and Codex. Watch our Head of AI Education @Prof_oz show you how: Ho

model-releasesfireworks-ai--x
23 Jun 2026
Model Releases

CAOA -- Completion-Assisted Object-CAD Alignment

DGX agent

arXiv:2606.18429v2 Announce Type: replace Abstract: Accurately aligning CAD models to their corresponding objects in indoor RGB-D scans is a central challenge in 3D semantic reconstruction. The task r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales

DGX agent

arXiv:2606.21949v1 Announce Type: new Abstract: Accurate and comprehensive video captions with consistent subject references are critical for downstream understanding and generation tasks. However, fe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Catching Lies Without Sending the Video: Privacy-Preserving Multimodal Deception Detection

DGX agent

arXiv:2606.22699v1 Announce Type: new Abstract: Frontier multimodal models can guess whether a person is lying from a testimony video. To do so, they stream that raw face and voice to a third-party mo

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CDER-SME: A Cross-Device Event-RGB Micro-Expression Dataset under Multi-Level Stress Induction

DGX agent

arXiv:2606.20715v1 Announce Type: new Abstract: Micro-expression recognition (MER) in realistic scenarios demands high temporal sensitivity and ecological validity, yet existing benchmarks are largely

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CFAgentBench: A Reproducible Environment and Benchmark for Autonomous Construction-Finance Agents

DGX agent

arXiv:2606.22000v1 Announce Type: cross Abstract: We introduce CFAgentBench, a reproducible, self-hostable environment and benchmark for autonomous construction-finance agents: a CFO/controller-class

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

DGX agent

arXiv:2606.22862v1 Announce Type: new Abstract: Forced chain-of-thought (CoT) is widely assumed to make vision-language models more reliable on video question answering. We propose a small three-probe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Chehre: An Emoji-Prompted Video Dataset for Perceptually Diverse Facial Expression Recognition

DGX agent

arXiv:2606.21657v1 Announce Type: new Abstract: Facial expressions are nonverbal social signals used in human interaction, but facial expression recognition datasets often focus on static images, basi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Chem2Gen-Bench: Benchmarking Chemical-to-Genetic Translation in Perturbation Response Space

DGX agent

arXiv:2606.21109v1 Announce Type: new Abstract: Virtual-cell and perturbation models are increasingly used to predict cellular responses for biomedical discovery, but chemical and genetic perturbation

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

CheXpercept: A Benchmark for Evaluating Expert-Level Lesion Perception in Chest X-rays

DGX agent

arXiv:2606.21020v1 Announce Type: new Abstract: The evaluation of vision-language models (VLMs) for chest X-ray (CXR) analysis has largely been limited to disease-presence classification without visua

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

ChronoLock: Protecting Videos from Unauthorized Text-to-Video Personalization

DGX agent

arXiv:2606.21146v1 Announce Type: new Abstract: Text-to-video (T2V) diffusion models have made it increasingly easy to synthesize realistic and temporally coherent videos, while recent personalization

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Claude is really proactive with Claude Tag. You don’t need to prompt it to do work, it can do work proactively based on your instructions. I…

DGX agent

Claude is really proactive with Claude Tag. You don’t need to prompt it to do work, it can do work proactively based on your instructions. It has excellent memory and access to your data, so it can be

model-releasesboris-cherny--x
23 Jun 2026
Model Releases

Claude Tag is an incredible new form factor for agents, so I think it's going to take some time to figure out the best practices, but these …

DGX agent

Claude Tag is an incredible new form factor for agents, so I think it's going to take some time to figure out the best practices, but these are some of my favorites 🧵 Introducing Claude Tag, a new way

model-releasesthariq--x
23 Jun 2026
Model Releases

Clipping the Price of Adaptivity at the Tail

DGX agent

arXiv:2606.22669v1 Announce Type: new Abstract: Adaptive stochastic convex optimization (SCO) methods face a fundamental ``price of adaptivity'' barrier: under the standard set of assumptions, they ca

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Cluster-Specific Localized Drift Detection for Efficient Batch Model Adaptation under Controlled Distribution Shift

DGX agent

arXiv:2606.22026v1 Announce Type: new Abstract: Machine learning systems deployed in dynamic environments frequently operate under nonstationary data distributions, where controlled distribution shift

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

CodePercept: Code-Grounded Visual STEM Perception for MLLMs

DGX agent

arXiv:2603.10757v2 Announce Type: replace Abstract: When MLLMs fail at Science, Technology, Engineering, and Mathematics (STEM) visual reasoning, a fundamental question arises: is it due to perceptual

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Comparative Evaluation of Machine Learning and Deep Learning Models for Wound-Rotor Synchronous Motor Performance Prediction

DGX agent

arXiv:2606.21230v1 Announce Type: new Abstract: Wound rotor synchronous motors have emerged as a strong alternative that eliminates dependence on REEs. However, WRSM design requires the simultaneous o

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Concept-Constrained Prompt Learning for Few-Shot CLIP Adaptation

DGX agent

arXiv:2606.22567v1 Announce Type: new Abstract: Few-shot prompt learning is an effective strategy for adapting CLIP to downstream tasks, but class-only prompt optimization can overfit base-class super

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Confidently Wrong: Severity-Aware Calibration of Prompt-Injection Detectors under Attack Shift

DGX agent

arXiv:2606.22659v1 Announce Type: cross Abstract: Prompt-injection detectors are deployed as guards: a model scores an input and a downstream system trusts or blocks it on that score. I study the conf

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ConnectomeBench2: A Unified Benchmark for Automated Connectomic Proofreading

DGX agent

arXiv:2606.21116v1 Announce Type: new Abstract: Proofreading--correcting segmentation errors in 3D brain reconstructions--is the rate-limiting step in synapse-resolution connectomics. We release Conne

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CoRDE: Concept-Prior Routed Diffusion Experts for Structural Generalization in Robot Manipulation

DGX agent

arXiv:2606.21935v1 Announce Type: new Abstract: Diffusion models excel at capturing multi-modal action distributions in robot imitation learning. However, in multi-task and long-horizon scenarios, mon

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

CoSA: Correlation-Guided Change Attention with Learnable Residual Gating for Remote Sensing Change Detection

DGX agent

arXiv:2606.21932v1 Announce Type: new Abstract: Remote sensing change detection (CD) from bi-temporal imagery is critical for applications such as urban monitoring, disaster assessment, and environmen

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CourseBlueprint: A Structured Pipeline for Adaptive Pedagogical Video Generation Grounded in Course Corpora

DGX agent

arXiv:2606.20608v1 Announce Type: cross Abstract: Generative text-to-video systems can produce visually fluent educational clips, but they rarely encode the pedagogical content knowledge (PCK) needed

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CRAX: Fast Safe Reinforcement Learning Benchmarking

DGX agent

arXiv:2606.20376v2 Announce Type: replace Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While bench

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model

DGX agent

arXiv:2606.22317v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is widely viewed as a promising path toward continuously improving large language models. Recent w

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Curvature-aware 3D length estimation of greenhouse cucumbers using RGB-D imaging and cubic spline arc-length integration

DGX agent

arXiv:2606.22439v1 Announce Type: new Abstract: Commercial greenhouse cucumber production is graded by fruit length, which drives harvest scheduling, labour allocation, and logistics. Manual measureme

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CVSBench: A Comprehensive Benchmark for Cross-view Spatial Reasoning and Dreaming

DGX agent

arXiv:2606.22476v1 Announce Type: new Abstract: Humans can effortlessly reason about scenes across different viewpoints, yet it remains unclear whether Vision-Language Models (VLMs) possess similar cr

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Data Pruning: Redundant, Problematic, and Interdependent Samples

DGX agent

arXiv:2606.21916v1 Announce Type: new Abstract: The performance of deep learning models is affected by not only data quantity but also data quality. Data pruning is a process by which practitioners ca

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams

DGX agent

arXiv:2606.21337v1 Announce Type: new Abstract: Massive unstructured multimodal streams suffer from high 'data entropy,' impeding both efficient human knowledge acquisition and high-quality AI post-tr

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Decoupling the Declarative from the Procedural in Vision-Language-Action Models

DGX agent

arXiv:2606.21496v1 Announce Type: cross Abstract: Deploying generalist robotic agents in the real world requires transferable skills. Specifically, a policy trained to clone a behavior from object-spe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Deep Learning-Based Sign Language Recognition from Videos and Cross-Lingual Translation to Indian Vernaculars

DGX agent

arXiv:2606.22494v1 Announce Type: cross Abstract: Sign language is a primary mode of communication for the global deaf and hard-of-hearing community, yet automated tools that recognize sign gestures f

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Demystifying Numerical Instability in LLM Inference: Achieving Reproducible Inference for Mission-Critical Tasks with HEAL

DGX agent

arXiv:2606.21023v1 Announce Type: new Abstract: As Large Language Models (LLMs) deploy into mission-critical domains (e.g., finance, medicine, and law), output reproducibility has become a strict syst

model-releasesarxiv-cs-lg
23 Jun 2026
← Previous
1…175176177178179…472
Next →