AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,611 results
Model Releases

What Do Language Priors Contribute to Darcy-Flow Inversion? A Mechanistic Audit

DGX agent

arXiv:2606.24967v1 Announce Type: new Abstract: In ill-posed inverse problems, the recovered solution depends as much on the prior as on the data, yet much of the engineering knowledge that could serv

model-releasesarxiv-cs-lg
25 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

What Does the Brain See? Multiview Neural Representations to Demystify the Brain-Visual Alignment

DGX agent

arXiv:2606.25718v1 Announce Type: new Abstract: Zero-shot visual decoding from electroencephalography (EEG) aims to infer visual semantics from non-invasive neural recordings, but remains challenging

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

What happens when Claude Code gets an experiment tracker

DGX agent

At CVPR 2026, Lambda ran a live demo for two and a half days: Claude Code teaching Google's Gemma 4 to play a Tetris-like game. Claude Code started with a Gemma 4 model that couldn't play at all. It p

model-releaseslambda-labs
25 Jun 2026
Model Releases

What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics

DGX agent

arXiv:2606.25182v1 Announce Type: new Abstract: Jailbreak attacks reveal a persistent weakness in aligned Large Language Models: carefully crafted prompts can elicit policy-violating responses despite

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

When Multi-Sensor Fusion Fails to Generalize: Cattle Posture Classification Under Animal-Level and Temporal Distribution Shift

DGX agent

arXiv:2606.24986v1 Announce Type: new Abstract: Automated cattle posture-classification systems frequently report near-perfect accuracy, yet their robustness under realistic deployment conditions rema

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

When you use Cohere, there are no staggered releases. No sudden disablements. We trust you completely: '[The customer] is in full control. W…

DGX agent

When you use Cohere, there are no staggered releases. No sudden disablements. We trust you completely: '[The customer] is in full control. We can't see in, we can't switch it off' - CEO @aidangomez Me

model-releasescohere--x
25 Jun 2026
Model Releases

While we eagerly await Fable 5's return, our agentic WebGPU kernel optimization framework kept running. Opus 4.8 picked up where Fable left …

DGX agent

While we eagerly await Fable 5's return, our agentic WebGPU kernel optimization framework kept running. Opus 4.8 picked up where Fable left off, pushing Liquid AI's new LFM2.5 230M to an unbelievable

model-releasesclem-delangue--x
25 Jun 2026
Model Releases

WOLF-VLA: Whole-Body Humanoid Optimal Locomotion Framework for Vision-Language-Action Learning

DGX agent

arXiv:2606.25591v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently demonstrated strong generalization in robotic manipulation, yet their applicability to whole-body, con

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

3/3 We took 7 weak agents (ranks 7-13, none scoring >45) from the leaderboard & merged them into 1 report/task. Essentially boosting for dee…

DGX agent

3/3 We took 7 weak agents (ranks 7-13, none scoring >45) from the leaderboard & merged them into 1 report/task. Essentially boosting for deep research. The result: New #1 DRB II TotalScore of 64.38. F

model-releasesai21-labs--x
24 Jun 2026
Model Releases

3DCarGen: Scalable 3D Car Generation via 3D-consistent Multi-view Synthesis

DGX agent

arXiv:2606.24257v1 Announce Type: new Abstract: High-quality 3D vehicle assets are essential for autonomous driving simulation. Although multi-view diffusion-based paradigms enable controllable single

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

5.2 could be better with more RL ...

DGX agent

5.2 could be better with more RL ... Deepswe's benchmark results are my own experience. I've used all models, GLM 5.2 ≈ Claude Opus 4.6–4.7. Kimi 2.7 code more like inference optimization. Looking for

model-releasesollama--x
24 Jun 2026
Model Releases

A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy

DGX agent

arXiv:2606.24115v1 Announce Type: cross Abstract: Vision-language models (VLMs) are prone to hallucination, which remains a major barrier to their safe deployment in clinical practice. To date, most h

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Benchmark of State-Space Models vs. Transformers and BiLSTM-based Models for Historical Newspaper OCR

DGX agent

arXiv:2604.00725v2 Announce Type: replace Abstract: End-to-end OCR for historical newspapers remains challenging, as models must handle long text sequences, degraded print quality, and complex layouts

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

A Hybrid Quantum-Classical Approach for Melt Pool Prediction in Laser Powder Bed Fusion

DGX agent

arXiv:2606.23719v1 Announce Type: cross Abstract: Laser powder bed fusion (LPBF) is a promising additive manufacturing technique that suffers from quality assurance concerns. Predicting melt pools fro

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

A Pairwise Human-Human Interaction Detection and Recognition Framework for Mobile Service Robots

DGX agent

arXiv:2602.22346v2 Announce Type: replace Abstract: Autonomous mobile service robots, such as lawnmowers or cleaning robots, operating in human-populated environments need to reason about human-human

model-releasesarxiv-cs-ro
24 Jun 2026
Model Releases

A Paninian Foundation for Indic Language Processing

DGX agent

arXiv:2606.24172v1 Announce Type: cross Abstract: More than a billion people communicate in Indic languages, yet the natural language processing infrastructure serving them remains fragmented and unde

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Physics-Informed Fourier-Wavelet Transformer for Multiscale Computational Fluid Dynamics Surrogate Modeling

DGX agent

arXiv:2606.24696v1 Announce Type: cross Abstract: Physics-informed surrogate models can accelerate computational fluid dynamics simulations. However, many existing methods reproduce global flow patter

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

A specialized reasoning large language model for accelerating rare disease diagnosis: a randomized AI physician assistance trial

DGX agent

arXiv:2606.24510v1 Announce Type: new Abstract: Rare diseases affect millions of individuals worldwide, yet timely diagnosis remains a major public health challenge due to scarcity of specialized clin

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Synthetic Reliability-Aware PINN Benchmark for Offshore Wind Turbine Support-Structure Monitoring with Bayesian Inverse Identification

DGX agent

arXiv:2606.24176v1 Announce Type: new Abstract: Reliable structural health monitoring (SHM) of offshore wind turbine (OWT) support structures requires fast state estimation from sparse measurements. R

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

ABACUS: Adapting Unified Foundation Model for Bridging Image Count Understanding and Generation

DGX agent

arXiv:2606.23835v1 Announce Type: new Abstract: ABACUS is a unified vision-language model that handles object counting, crowd counting, referring-expression counting, and count-faithful image generati

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Acquisition state behaves as a structured, measurable variable governing lung-nodule AI: kernel-driven measurement instability and noise-driven detection fragility, invisible to DICOM metadata

DGX agent

arXiv:2606.12824v2 Announce Type: replace-cross Abstract: AI governance for medical imaging is formalizing: the 2026 ACR-SIIM Practice Parameter recommends local acceptance testing and ongoing drift m

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

DGX agent

arXiv:2606.24589v1 Announce Type: new Abstract: Scaling adversarial evaluation of large language models requires both a method for generating hard inputs and a reliable way to confirm that resulting f

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War

DGX agent

arXiv:2606.24391v1 Announce Type: new Abstract: We introduce Age of LLM, a turn-based 1v1 benchmark in which two LLMs face off on a 13x7 grid to destroy the enemy base. Three stressors are deliberate:

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

AI-Driven Predictive Maintenance with Environmental Context Integration for Connected Vehicles: Simulation, Benchmarking, and Field Validation

DGX agent

arXiv:2603.13343v3 Announce Type: replace-cross Abstract: Predictive maintenance for connected vehicles offers the potential to reduce unexpected breakdowns and improve fleet reliability, but most exi

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approach

DGX agent

arXiv:2606.24655v1 Announce Type: cross Abstract: The explosive growth and complexity of product data within the dynamic Brazilian e-commerce landscape demand robust and specialized methods for struct

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

[AINews] Claude Tag: Multiplayer, Proactive, Persistent Agents in Slack

DGX agent

Anthropic's Claude API now supports multiplayer agent capabilities that enable multiple Claude instances to collaborate within Slack, with features for proactive behavior and persistent state manageme

model-releaseslatent-space
24 Jun 2026
Model Releases

An analysis of GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Gab's Arya, and other AI models: most chatbots frequently provide left-leaning responses to political prompts (Kevin Schaul/Washington Post)

DGX agent

Kevin Schaul / Washington Post: An analysis of GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Gab's Arya, and other AI models: most chatbots frequently provide left-leaning responses to political prompts — Excerp

model-releasestechmeme
24 Jun 2026
Model Releases

Anthropic debuts Claude Tag, a more capable AI teammate that lives within Slack

DGX agent

Anthropic PBC today unveiled a new version of its chatbot Claude that lives inside Slack, where it operates like a virtual employee. It’s called Claude Tag, and it’s designed to work across entire org

model-releasessiliconangle
24 Jun 2026
Model Releases

Are Text-to-Image Models Inductivist Turkeys? A Counterfactual Benchmark for Causal Reasoning

DGX agent

arXiv:2606.24548v1 Announce Type: new Abstract: Text-to-image (T2I) generation models have achieved remarkable progress in producing visually realistic images from natural language prompts. Yet it rem

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Are We Ready For An Agent-Native Memory System?

DGX agent

arXiv:2606.24775v1 Announce Type: new Abstract: Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.24601v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) addresses the problem of training multiple agents that pursue collaborative, competitive, or mixed objectives.

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Assessing Distribution Shift in Human Activity Recognition for Domain Generalization

DGX agent

arXiv:2606.24781v1 Announce Type: new Abstract: While the field of Human Activity Recognition (HAR) continues to draw interest from researchers and advance in important ways, some key challenges remai

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

At Hugging Face we've been building our own agent that we use via Slack (Moon Bot). Honestly, building your own is quite simple and you'll b…

DGX agent

At Hugging Face we've been building our own agent that we use via Slack (Moon Bot). Honestly, building your own is quite simple and you'll be happy you did: any model you want (self-hosted if needed),

model-releasesclem-delangue--x
24 Jun 2026
Model Releases

AutoSpecNER: A Fine-Grained Named Entity Recognition Dataset for Vehicle Specification Extraction

DGX agent

arXiv:2606.24387v1 Announce Type: new Abstract: Vehicle advertisements contain rich specification information, but automotive NER resources remain limited. We introduce AutoSpecNER, an expert-annotate

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Average Rankings Mask Per-Subject Optimality: A Friedman-Nemenyi Benchmark of EEG Motor-Imagery BCI Decoders

DGX agent

arXiv:2606.24394v1 Announce Type: cross Abstract: Electroencephalography (EEG) is the dominant non-invasive modality for brain-computer interfaces (BCIs), yet reliable decoding of motor imagery is ham

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

awesome launch

DGX agent

awesome launch We built Claude for outbound sellers. AEs & SDRs can harness GTM engineering through chat across 40+ data sources, no technical skills required. We’ve had 57,548 queries in our first fe

model-releasesharrison-chase--x
24 Jun 2026
Model Releases

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks

DGX agent

arXiv:2606.24162v1 Announce Type: new Abstract: Foundation models have been increasingly applied to behavioral science domains such as psychology, sociology, and economics. While these models show pro

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions

DGX agent

arXiv:2501.11790v5 Announce Type: replace-cross Abstract: Recent studies have raised significant concerns regarding the reliability of current mathematics benchmarks, highlighting issues such as simpl

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BenchX: Benchmarking AI Models for Cancer Detection and Localization with Demographic and Protocol Biases

DGX agent

arXiv:2606.24883v1 Announce Type: new Abstract: Artificial intelligence (AI) has achieved remarkable success in medical imaging, but it is widely recognized that these models often perform inconsisten

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Big Tech's $1 trillion AI moat just got DESTROYED by a free Chinese download. Microsoft, Amazon, Google, and Meta are pouring fortunes into …

DGX agent

Big Tech's 1 trillion AI moat just got DESTROYED by a free Chinese download. Microsoft, Amazon, Google, and Meta are pouring fortunes into chips and data centers because they have been told that whoev

model-releasesgary-marcus--x
24 Jun 2026
Model Releases

BIM-Edit: Benchmarking Large Language Models for IFC-Based Building Information Modeling

DGX agent

arXiv:2606.20146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to computer-aided design (CAD) to generate design artifacts from textual instructions. In engi

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BioMedArena: An Open-source Toolkit for Building and Evaluating Biomedical Deep Research Agents

DGX agent

arXiv:2605.06177v2 Announce Type: replace Abstract: Reproducing and comparing deep research agents today is hard: the same backbone evaluated on the same benchmark can report different accuracies acro

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BioMedVR: Confusion-Aware Mixture-of-Prompt Experts for Biomedical Visual Reprogramming

DGX agent

arXiv:2606.24740v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) such as CLIP have demonstrated strong generalization across natural-image domains. However, adapting th

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Blockwise Policy-Drift Gating for On-Policy Distillation

DGX agent

arXiv:2606.24084v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student policy using teacher signals computed on trajectories sampled by the student itself. Recent work shows t

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BluTrain: A C++/CUDA Framework for AI Systems

DGX agent

arXiv:2606.24780v1 Announce Type: new Abstract: Progress in deep learning is, at scale, more a matter of systems engineering than of modelling: the behaviour of a model in training (its throughput, it

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's. now they have beat deepseek with t…

DGX agent

btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's. now they have beat deepseek with the world's undisputed top open model and in some respects (s

model-releasesswyx--x
24 Jun 2026
Model Releases

Business as Rulesual: A Benchmark and Framework for Business Rule Flow Modeling with LLMs

DGX agent

arXiv:2505.18542v4 Announce Type: replace Abstract: Extracting structured procedural knowledge from unstructured business documents is a critical yet unresolved bottleneck in process automation. While

model-releasesarxiv-cs-cl
24 Jun 2026
← Previous
1…169170171172173…472
Next →