AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,778 results
Model Releases

Hybrid Compression: Integrating Pruning and Quantization for Optimized Neural Networks

DGX agent

arXiv:2606.22935v1 Announce Type: new Abstract: Deep neural networks have witnessed remarkable advancements in recent years and have become integral to various applications. However, alongside these d

model-releasesarxiv-cs-cv
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

HyperQuant: A Rate-Distortion-Optimal Quantization Pipeline for Large Language and Diffusion Models

DGX agent

arXiv:2606.23406v1 Announce Type: new Abstract: We present HyperQuant (Hadamard, optimallY Packing, Entropy Rice-coding), a unified post-training quantization pipeline for the weights and the KV cache

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

I also have Claude monitoring Slack channels to proactively respond to people to answer their questions and draft PRs, and to react with spe…

DGX agent

I also have Claude monitoring Slack channels to proactively respond to people to answer their questions and draft PRs, and to react with specific emojis like ✅❌ when a thread is resolved. To set it up

model-releasesboris-cherny--x
23 Jun 2026
Model Releases

I guarantee you are sleeping on small models. Deepseek V4 Flash can do ~80% of the tasks you ask Claude or Codex for. It is 137x cheaper per…

DGX agent

DeepSeek V4 Flash is a small language model capable of handling approximately 80% of common tasks performed by larger models like Claude or Codex, while offering significantly lower costs (137x cheape

model-releasesclem-delangue--x
23 Jun 2026
Model Releases

I tag Claude many times a day to write PRs, address user feedback, investigate incidents, do data analyses, summarize information, etc. Use …

DGX agent

I tag Claude many times a day to write PRs, address user feedback, investigate incidents, do data analyses, summarize information, etc. Use it as your company's search engine. 'What's the status on X?

model-releasesboris-cherny--x
23 Jun 2026
Model Releases

IMAGIN-4D: Image-Guided Controllable Interaction Generation

DGX agent

arXiv:2606.23675v1 Announce Type: new Abstract: Generating human-object interactions (HOI) is central to character animation, robotics, AR/VR, and embodied AI. Recent HOI generation methods synthesize

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Improving Text-to-Music Generation with Human Preference Rewards

DGX agent

arXiv:2606.21670v1 Announce Type: cross Abstract: We describe our entry to the efficiency track of the Academic Text-to-Music (ATTM) Grand Challenge at ICME 2026. Beyond the challenge protocol's FAD-C

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

In-Context Molecular Property Prediction with LLMs: A Blinding Study on Memorization and Knowledge Conflicts

DGX agent

arXiv:2603.25857v2 Announce Type: replace Abstract: The capabilities of large language models (LLMs) have expanded beyond natural language processing to scientific prediction tasks, including molecula

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

In LLM Reasoning, there is Irrationality on top of Value Misalignment

DGX agent

arXiv:2606.20624v1 Announce Type: cross Abstract: Significant progress has been made in aligning LLMs with target value functions. We argue that, even when an LLM has been well aligned in (post-)train

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

IndicGuard: A Multilingual Safety Guard Model and Dataset for Indic Languages

DGX agent

arXiv:2606.22841v1 Announce Type: cross Abstract: As Large Language Models (LLMs) achieve widespread integration across diverse linguistic landscapes, ensuring their safety and alignment with regional

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Intend, Reflect, Refine: An Adaptive Multimodal Reflection Framework for Autonomous Driving

DGX agent

arXiv:2606.22913v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving by incorporating reasoning for better interpretability and planni

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Introducing Claude code for 3D modeling You can now create a production-level 3D models from any agent you are using. Ready to plug into you…

DGX agent

Introducing Claude code for 3D modeling You can now create a production-level 3D models from any agent you are using. Ready to plug into your own apps or games Made possible by @trysuzanne and @monid_

model-releasesyohei-nakajima--x
23 Jun 2026
Model Releases

Introducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and too…

DGX agent

Introducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and tools you choose. Tag Claude in and delegate tasks to it while

model-releasesyohei-nakajima--x
23 Jun 2026
Model Releases

Introducing Mistral OCR 4. It creates structure with bounding boxes, block classification, and inline confidence scores in 170 languages. 🧵…

DGX agent

Mistral OCR 4 is an optical character recognition model that extracts text with structural information including bounding boxes and block classification across 170 languages. The system provides inlin

model-releasesarthur-mensch--x
23 Jun 2026
Model Releases

Inverse Problem for Partial Differential Equations with Jump Discontinuities in Coefficients by Two-stage Physics-Informed Deep Learning and Statistical Mixture Models

DGX agent

arXiv:2510.14656v2 Announce Type: replace-cross Abstract: This work proposes a two-stage physics-informed deep learning framework that combines neural-network-based sampling with statistical inference

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

IOI: Decoupling Kinematics and Physics for Interactive World Models

DGX agent

arXiv:2606.23296v1 Announce Type: new Abstract: Developing generalist embodied agents requires interactive environments providing visually realistic feedback and accurate action-conditioned dynamics.

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Is Our Benchmark Enough? An Analysis of Continual Learning for MLLMs

DGX agent

arXiv:2606.20961v1 Announce Type: new Abstract: Continual adaptation is essential for multimodal large language models (MLLMs) deployed across evolving domains, but the state-of-the-art MR-LoRA method

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

It's Much Easier for Neural Networks to learn Game of Life Dynamics with the Right Activation Function: Polynomial Kolmogorov-Arnold Networks

DGX agent

arXiv:2606.23587v1 Announce Type: new Abstract: Previous work has found a gap between the scale of neural networks that reliably learn Conway's Game of Life, and minimal networks capable of representi

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Jury Duty: Calibration and Orientation Failures in MLLM-as-a-Judge Under Cultural Ambiguity

DGX agent

arXiv:2606.20676v1 Announce Type: new Abstract: MLLM-as-a-Judge is conventionally validated by agreement with human annotations, but this metric is undefined when the human pool is culturally heteroge

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Kamera: Unified Position-Invariant Multimodal KV Cache for Training-Free Reuse

DGX agent

arXiv:2606.23581v1 Announce Type: cross Abstract: Multimodal agents repeatedly re-examine the same video frames, UI screenshots, and rendered artifacts as their context window slides and reasoning ite

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

L20-Edu-135M: An Auditable Single-GPU Study of Data-Efficient Small Language Modeling

DGX agent

arXiv:2606.22189v1 Announce Type: new Abstract: Small language models are cheap to serve and feasible on local hardware, but strong public 135M-class systems are commonly trained with hundreds of bill

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Large Language Model-Assisted Cleaning of Report-Derived Labels in a Large-Scale Chest CT Dataset

DGX agent

arXiv:2606.22382v1 Announce Type: cross Abstract: Purpose: To evaluate whether large language model (LLM)-assisted label cleaning can identify label-report discordance in CT-RATE, a large-scale public

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

LAYUP: Asynchronous decentralized gradient descent with LAYer-wise UPdates

DGX agent

arXiv:2410.05985v4 Announce Type: replace Abstract: The increasing size of deep learning models has made distributed training across multiple devices essential. Synchronous, centralized methods incur

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning a Normal World Model for Few-Shot Boundary-Calibrated Abnormality Detection

DGX agent

arXiv:2606.22261v1 Announce Type: new Abstract: Abnormality detection in complex systems faces two practical barriers: abnormal labels are scarce, and binary labels do not quantify how far an event ha

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning-Augmented Algorithms for Online Vertex Cover

DGX agent

arXiv:2606.22831v1 Announce Type: cross Abstract: This paper studies learning-augmented online weighted vertex cover with advice and a parameter lambda in (0,1). We consider two graph cases: bipartite

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning Bug Context for PyTorch-to-JAX Translation with LLMs

DGX agent

arXiv:2510.09898v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on code translation between widely used programming languages. However, translation becom

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning by Shifting: Temporal View Construction for Time Series Contrastive Learning

DGX agent

arXiv:2606.21957v1 Announce Type: new Abstract: Supervised learning demands large quantities of labeled data, a bottleneck that is expensive and reliant on domain-specific expertise. Self-supervised l

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Leaving aside the “Vandalism” stuff for a brief moment: President Trump’s claims today that the reflecting pool “never opened” or was “rarel…

DGX agent

Leaving aside the “Vandalism” stuff for a brief moment: President Trump’s claims today that the reflecting pool “never opened” or was “rarely open” after the Obama repair project are just not true. It

model-releasesanthropic--x
23 Jun 2026
Model Releases

Leveraging LaBSE with Progressive Curriculum Learning for Multicultural Polarization

DGX agent

arXiv:2606.21718v1 Announce Type: cross Abstract: Detecting online polarization remains a critical challenge, particularly in multilingual and multicultural contexts where intergroup hostility is prev

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

LIBERO-Safety: A Comprehensive Benchmark for Physical and Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2606.23686v1 Announce Type: new Abstract: Despite the impressive manipulation capabilities of Vision-Language-Action (VLA) models, their operational safety under strict constraints remains large

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

LLM-Based Generalizable Hierarchical Task Planning and Execution for Heterogeneous Robot Teams with Event-Driven Replanning

DGX agent

arXiv:2511.22354v2 Announce Type: replace Abstract: This paper introduces CoMuRoS (Collaborative Multi-Robot System), a generalizable hierarchical architecture for heterogeneous robot teams that unifi

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Load Testing for Machine Learning Model Serving Systems at Scale

DGX agent

arXiv:2606.22013v1 Announce Type: new Abstract: Machine learning (ML) model serving has become a dominant consumer of GPU infrastructure, yet capacity planning in these systems remains largely ad hoc.

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Localizing and Editing Knowledge in Large Audio-Language Models

DGX agent

arXiv:2603.14343v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) have shown strong performance in speech understanding, making speech a natural interface for accessing factual i

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

LoCC: Detection and Localization of Lip-Syncing Deepfakes via Counterfactual Frame Consistency

DGX agent

arXiv:2606.22772v1 Announce Type: new Abstract: Lip-syncing deepfakes are among the most challenging forms of manipulated media because their artifacts are localized almost exclusively to the mouth re

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

LOGOS: LiDAR-Only Gaussian Elevation Splatting for Unified Tiny Obstacle Segmentation

DGX agent

arXiv:2606.21527v1 Announce Type: cross Abstract: Robust obstacle segmentation is essential for the safety of intelligent robots, where LiDAR-based perception systems play a fundamental role in the ro

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

(Look mom - I made it!) Learn more about configuring agent identity and some of the key decisions we made for Claude Tag- https://claude.com…

DGX agent

This post likely discusses configuration options for setting up Claude agent identity and explains the design decisions behind Claude Tag, Anthropic's tool for customizing Claude's behavior and person

model-releasesthariq--x
23 Jun 2026
Model Releases

Low-variance estimators overcome the phase-gradient bottleneck in complex-valued neural quantum states

DGX agent

arXiv:2606.13912v2 Announce Type: replace-cross Abstract: Complex neural quantum states are difficult to optimize when their wavefunction phase carries gauge, chiral, fermionic, or topological structu

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

LUMINA-26: Low-Light Understanding for Modeling and Interpreting Night-time Actions

DGX agent

arXiv:2606.23118v1 Announce Type: new Abstract: Low-light human action recognition remains a challenging problem due to poor illumination, amplified noise, motion ambiguity, and diverse real-world sce

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models

DGX agent

arXiv:2509.23729v3 Announce Type: replace Abstract: Large Language Models (LLMs) with multimodal capabilities have revolutionized vision-language tasks, but their deployment often requires huge memory

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

MammoExpert: Benchmarking Chain-of-Thought Reasoning in Mammography Diagnosis

DGX agent

arXiv:2606.21119v1 Announce Type: new Abstract: Mammography is an essential tool for breast cancer detection, with millions of examinations conducted annually. However, publicly available high-quality

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

MapReason-OSM: Can Vision-Language Models Make Graph-Verifiable Mobility Decisions from Street Maps ?

DGX agent

arXiv:2606.22597v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to read maps for logistics, delivery, and accessible navigation, where the output is an actionable d

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Mat-Pref: Verifiable-Reward Training Improves Compositional Reasoning in Inorganic Materials

DGX agent

arXiv:2606.21830v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has driven rapid progress in mathematical and code reasoning, but when extended to science, existi

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Measuring Intent Comprehension in LLMs

DGX agent

arXiv:2506.16584v3 Announce Type: replace-cross Abstract: People judge interactions with large language models (LLMs) as successful when outputs match what they want, not what they type. Yet LLMs are

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

MEDLAYXPLAIN: Benchmarking the Expert-Lay Gap in Medical Vision-Language Models

DGX agent

arXiv:2606.21194v1 Announce Type: new Abstract: Medical Vision-Language Models (Med-VLMs) achieve strong expert-level performance, yet their ability to generate patient-accessible descriptions remains

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

MedTS-TTT: Test-Time Training for Medical Time Series Classification

DGX agent

arXiv:2606.21329v1 Announce Type: new Abstract: Medical time series (MedTS) signals such as electroencephalography (EEG) and electrocardiography (ECG) support many clinical applications. However, subs

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Mesh2GS: White-Box 3DGS Construction via Plenoptic Sampling

DGX agent

arXiv:2606.21898v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a promising method for high-quality, real-time 3D reconstruction. To associate 3DGS with mesh representati

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Meta launches cheaper smart glasses without Ray-Ban

DGX agent

For the past three years, 'Meta' and 'Ray-Ban' have been synonymous in the smart glasses space. Not anymore. Yesterday, I slipped on several pairs of Meta Glasses - no Ray-Bans - in three different st

model-releasesthe-verge-ai
23 Jun 2026
Model Releases

Mirage: a Clean-Label Backdoor against LiDAR 3D Object Detection

DGX agent

arXiv:2606.20752v1 Announce Type: new Abstract: Deep neural network-based LiDAR 3D object detection serves as a critical perception component in safety-critical autonomous systems. However, recent stu

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…181182183184185…475
Next →