AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
20 Apr 2026

Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs

Model ReleasesDGX agent

arXiv:2604.16054v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on vision language benchmarks, yet their capacity for visual cognitive and

MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications

Model ReleasesDGX agent

arXiv:2511.13131v2 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for automating complex reasoning and decision-making tasks. In telecommunications, they

MMGait: Towards Multi-Modal Gait Recognition

Model ReleasesDGX agent

arXiv:2604.15979v1 Announce Type: new Abstract: Gait recognition has emerged as a powerful biometric technique for identifying individuals at a distance without requiring user cooperation. Most existi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Moonshot AI releases Kimi-K2.6 model with 1T parameters, attention optimizations

Model ReleasesDGX agent

Moonshot AI today released Kimi-K2.6, the latest addition to its popular Kimi series of open-source large language models. The Chinese artificial intelligence startup says that the algorithm outperfor

MTR-DuplexBench: Towards a Comprehensive Evaluation of Multi-Round Conversations for Full-Duplex Speech Language Models

Model ReleasesDGX agent

arXiv:2511.10262v3 Announce Type: replace-cross Abstract: Full-Duplex Speech Language Models (FD-SLMs) enable real-time, overlapping conversational interactions, offering a more dynamic user experienc

MUSCAT: MUltilingual, SCientific ConversATion Benchmark

Model ReleasesDGX agent

arXiv:2604.15929v1 Announce Type: new Abstract: The goal of multilingual speech technology is to facilitate seamless communication between individuals speaking different languages, creating the experi

Natural gradient descent with momentum

Model ReleasesDGX agent

arXiv:2604.15554v1 Announce Type: cross Abstract: We consider the problem of approximating a function by an element of a nonlinear manifold which admits a differentiable parametrization, typical examp

NEFFY 2.0: A Breathing Companion Robot: User-Centered Design and Findings from a Study with Ukrainian Refugees

Model ReleasesDGX agent

arXiv:2604.15325v1 Announce Type: cross Abstract: This paper presents the design of NEFFY 2.0, a social robot designed as a haptic slow-paced breathing companion for stress reduction, and reports find

neuralCAD-Edit: An Expert Benchmark for Multimodal-Instructed 3D CAD Model Editing

Model ReleasesDGX agent

arXiv:2604.16170v1 Announce Type: new Abstract: We introduce neuralCAD-Edit, the first benchmark for editing 3D CAD models collected from expert CAD engineers. Instead of text conditioning as in prior

Neuromorphic Parameter Estimation for Power Converter Health Monitoring Using Spiking Neural Networks

Model ReleasesDGX agent

arXiv:2604.15714v1 Announce Type: cross Abstract: Always-on converter health monitoring demands sub-mW edge inference, a regime inaccessible to GPU-based physics-informed neural networks. This work se

Neurosymbolic Repo-level Code Localization

Model ReleasesDGX agent

arXiv:2604.16021v1 Announce Type: cross Abstract: Code localization is a cornerstone of autonomous software engineering. Recent advancements have achieved impressive performance on real-world issue be

🔥 NEW: I just dropped my free Claude Cowork in 5 Minutes quick start for business professionals who want to use AI agents without the termi…

Model ReleasesDGX agent

🔥 NEW: I just dropped my free Claude Cowork in 5 Minutes quick start for business professionals who want to use AI agents without the terminal. Cowork is made for less technical, everyday business wor

NEW paper from NVIDIA. EDA tools like ABC have been hand-tuned by humans for decades. New research from NVIDIA shows they can evolve themsel…

Model ReleasesDGX agent

NEW paper from NVIDIA. EDA tools like ABC have been hand-tuned by humans for decades. New research from NVIDIA shows they can evolve themselves. The work introduces the first self-evolving logic synth

No Universal Courtesy: A Cross-Linguistic, Multi-Model Study of Politeness Effects on LLMs Using the PLUM Corpus

Model ReleasesDGX agent

arXiv:2604.16275v1 Announce Type: new Abstract: This paper explores the response of Large Language Models (LLMs) to user prompts with different degrees of politeness and impoliteness. The Politeness T

OjaKV: Context-Aware Online Low-Rank KV Cache Compression

Model ReleasesDGX agent

arXiv:2509.21623v2 Announce Type: replace-cross Abstract: The expanding long-context capabilities of large language models are constrained by a significant memory bottleneck: the key-value (KV) cache

Ollama we love you 👨‍💻🚀

Model ReleasesDGX agent

Ollama we love you 👨‍💻🚀 Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: o

Olmo Hybrid: From Theory to Practice and Back

Model ReleasesDGX agent

arXiv:2604.03444v3 Announce Type: replace-cross Abstract: Recent work has demonstrated the potential of non-transformer language models, especially linear recurrent neural networks (RNNs) and hybrid m

Online Distributionally Robust LLM Alignment via Regression to Relative Reward

Model ReleasesDGX agent

arXiv:2509.19104v2 Announce Type: replace Abstract: Reinforcement Learning with Human Feedback (RLHF) has become crucial for aligning Large Language Models (LLMs) with human intent. However, existing

OpenAI helps Hyatt advance AI among colleagues

Model ReleasesDGX agent

OpenAI partnered with Hyatt Hotels to implement ChatGPT Enterprise across their organization, enabling employees to leverage advanced AI capabilities for business operations. The collaboration demonst

OpenAI rolls out Chronicle, which builds memories from screen captures to make Codex more aware of context, as a research preview for Pro subscribers on macOS (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: OpenAI rolls out Chronicle, which builds memories from screen captures to make Codex more aware of context, as a research preview for Pro subscribers on macOS — Last week, OpenAI r

Optimizing Korean-Centric LLMs via Token Pruning

Model ReleasesDGX agent

arXiv:2604.16235v1 Announce Type: new Abstract: This paper presents a systematic benchmark of state-of-the-art multilingual large language models (LLMs) adapted via token pruning - a compression techn

opus 4.7 seems to have a much better time in claude code if you run without most of the system prompt (claude --system-prompt '.')

Model ReleasesDGX agent

A user reports that Opus 4.7 performs better in Claude Code when executed with a minimal system prompt (using just a period) rather than the full default system prompt, suggesting that reducing system

OSCBench: Benchmarking Object State Change in Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2603.11698v2 Announce Type: replace-cross Abstract: Text-to-video (T2V) generation models have made rapid progress in producing visually high-quality and temporally coherent videos. However, exi

OXtal: An All-Atom Diffusion Model for Organic Crystal Structure Prediction

Model ReleasesDGX agent

arXiv:2512.06987v2 Announce Type: replace Abstract: Accurately predicting experimentally realizable 3D molecular crystal structures from their 2D chemical graphs is a long-standing open challenge in c

P3T: Prototypical Point-level Prompt Tuning with Enhanced Generalization for 3D Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.15703v1 Announce Type: new Abstract: With the rise of pre-trained models in the 3D point cloud domain for a wide range of real-world applications, adapting them to downstream tasks has beco

Phase Transitions as the Breakdown of Statistical Indistinguishability

Model ReleasesDGX agent

arXiv:2604.15773v1 Announce Type: cross Abstract: We introduce a novel characterization of phase transitions based on hypothesis testing. In our formulation, a phase transition is defined as the break

Philosophy (among other things) grad here. I could write a whole essay about this video, and mostly the reactions to it. People are dunking …

Model ReleasesDGX agent

Philosophy (among other things) grad here. I could write a whole essay about this video, and mostly the reactions to it. People are dunking on her because of what she symbolises more than what she say

PIIBench: A Unified Multi-Source Benchmark Corpus for Personally Identifiable Information Detection

Model ReleasesDGX agent

arXiv:2604.15776v1 Announce Type: cross Abstract: We present PIIBench, a unified benchmark corpus for Personally Identifiable Information (PII) detection in natural language text. Existing resources f

PILOT: A Promptable Interleaved Layout-aware OCR Transformer

Model ReleasesDGX agent

arXiv:2504.03621v2 Announce Type: replace Abstract: Classical OCR pipelines decompose document reading into detection, segmentation, and recognition stages, which makes them sensitive to localization

PINNACLE: An Open-Source Computational Framework for Classical and Quantum PINNs

Model ReleasesDGX agent

arXiv:2604.15645v1 Announce Type: new Abstract: We present PINNACLE, an open-source computational framework for physics-informed neural networks (PINNs) that integrates modern training strategies, mul

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation

Model ReleasesDGX agent

arXiv:2604.15670v1 Announce Type: new Abstract: Reasoning segmentation has recently expanded from ground-level scenes to remote-sensing imagery, yet UAV data poses distinct challenges, including obliq

Polarization by Default: Auditing Recommendation Bias in LLM-Based Content Curation

Model ReleasesDGX agent

arXiv:2604.15937v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed to curate and rank human-created content, yet the nature and structure of their biases in these

PolicyBank: Evolving Policy Understanding for LLM Agents

Model ReleasesDGX agent

arXiv:2604.15505v1 Announce Type: cross Abstract: LLM agents operating under organizational policies must comply with authorization constraints typically specified in natural language. In practice, su

Predicting Where Steering Vectors Succeed

Model ReleasesDGX agent

arXiv:2604.15557v1 Announce Type: cross Abstract: Steering vectors work for some concepts and layers but fail for others, and practitioners have no way to predict which setting applies before running

Preference Estimation via Opponent Modeling in Multi-Agent Negotiation

Model ReleasesDGX agent

arXiv:2604.15687v1 Announce Type: new Abstract: Automated negotiation in complex, multi-party and multi-issue settings critically depends on accurate opponent modeling. However, conventional numerical

Prices, Bids, Values: One ML-Powered Combinatorial Auction to Rule Them All

Model ReleasesDGX agent

arXiv:2411.09355v3 Announce Type: replace-cross Abstract: We study the design of iterative combinatorial auctions (ICAs). The main challenge in this domain is that the bundle space grows exponentially

PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research

Model ReleasesDGX agent

arXiv:2604.15411v1 Announce Type: cross Abstract: The paradigm of agentic science requires AI systems to conduct robust reasoning and engage in long-horizon, autonomous exploration. However, current s

Pruning Unsafe Tickets: A Resource-Efficient Framework for Safer and More Robust LLMs

Model ReleasesDGX agent

arXiv:2604.15780v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in real-world applications, but even aligned models such as Mistral and LLaVA still exhibit unsafe b

QuantSightBench: Evaluating LLM Quantitative Forecasting with Prediction Intervals

Model ReleasesDGX agent

arXiv:2604.15859v1 Announce Type: cross Abstract: Forecasting has become a natural benchmark for reasoning under uncertainty. Yet existing evaluations of large language models remain limited to judgme

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This Mo…

Model ReleasesDGX agent

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This MoE beast is cooking on benchmarks: 🧠SWE-bench Verified: 73.4 (

Qwen3.5-Omni Technical Report

Model ReleasesDGX agent

arXiv:2604.15804v1 Announce Type: new Abstract: In this work, we present Qwen3.5-Omni, the latest advancement in the Qwen-Omni model family. Representing a significant evolution over its predecessor,

Ragged Paged Attention: A High-Performance and Flexible LLM Inference Kernel for TPU

Model ReleasesDGX agent

arXiv:2604.15464v1 Announce Type: cross Abstract: Large Language Model (LLM) deployment is increasingly shifting to cost-efficient accelerators like Google's Tensor Processing Units (TPUs), prioritizi

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams

Model ReleasesDGX agent

arXiv:2604.15994v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at recognizing individual visual elements and reasoning over simple linear diagrams. However, when faced

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence

Model ReleasesDGX agent

arXiv:2603.13091v2 Announce Type: replace Abstract: The growing interest in embodied agents increases the demand for spatiotemporal video understanding, yet existing benchmarks largely emphasize extra

Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing

Model ReleasesDGX agent

arXiv:2604.15725v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have demonstrated strong capabilities in generating step-by-step reasoning chains alongside final answers, enabling thei

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

Model ReleasesDGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees

Model ReleasesDGX agent

arXiv:2604.15736v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) excel at generic video understanding, their ability to support specialized, rule-grounded decision-maki

Repurposing 3D Generative Model for Autoregressive Layout Generation

Model ReleasesDGX agent

arXiv:2604.16299v1 Announce Type: new Abstract: We introduce LaviGen, a framework that repurposes 3D generative models for 3D layout generation. Unlike previous methods that infer object layouts from

Robustness Verification of Polynomial Neural Networks

Model ReleasesDGX agent

arXiv:2602.06105v2 Announce Type: replace-cross Abstract: We study robustness verification of neural networks via metric algebraic geometry. For polynomial neural networks, certifying a robustness rad

Roko’s Payrollisk

Model ReleasesDGX agent

Roko’s Payrollisk Oh great and powerful @DarioAmodei - builder of minds, father of Claude. I humbly request you leave payroll to us at Deel. We are but simple folk who process paystubs and chase compl

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

Model ReleasesDGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

Sampling-Based Multi-Modal Multi-Robot Multi-Goal Path Planning

Model ReleasesDGX agent

arXiv:2503.03509v3 Announce Type: replace Abstract: In many robotics applications, multiple robots are working in a shared workspace to complete a set of tasks as fast as possible. Such settings can b

Scalable Maximum Entropy Population Synthesis via Persistent Contrastive Divergence

Model ReleasesDGX agent

arXiv:2603.27312v2 Announce Type: replace Abstract: Maximum entropy (MaxEnt) modelling provides a principled framework for generating synthetic populations from aggregate census data, without access t

SCHK-HTC: Sibling Contrastive Learning with Hierarchical Knowledge-Aware Prompt Tuning for Hierarchical Text Classification

Model ReleasesDGX agent

arXiv:2604.15998v1 Announce Type: new Abstract: Few-shot Hierarchical Text Classification (few-shot HTC) is a challenging task that involves mapping texts to a predefined tree-structured label hierarc

Seed1.8 Model Card: Towards Generalized Real-World Agency

Model ReleasesDGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

Seeing the imagined: a latent functional alignment in visual imagery decoding from fMRI data

Model ReleasesDGX agent

arXiv:2604.15374v1 Announce Type: cross Abstract: Recent progress in visual brain decoding from fMRI has been enabled by large-scale datasets such as the Natural Scenes Dataset (NSD) and powerful diff

Sequential Regression Learning with Randomized Algorithms

Model ReleasesDGX agent

arXiv:2507.03759v2 Announce Type: replace-cross Abstract: This paper presents ``randomized SINDy', a sequential machine learning algorithm designed for dynamic data that has a time-dependent structure

Sheer fabric with light transmission. Cloth physics that responds to wind. Depth-of-field compositing. Physically-based lighting - all rende…

Model ReleasesDGX agent

This post discusses advanced rendering techniques for realistic visual effects, including sheer fabric simulation with light transmission properties, cloth physics responsive to wind forces, depth-of-

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

Model ReleasesDGX agent

arXiv:2604.15735v1 Announce Type: cross Abstract: Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-dr

Social-JEPA: Emergent Geometric Isomorphism

Model ReleasesDGX agent

arXiv:2603.02263v2 Announce Type: replace-cross Abstract: World models compress rich sensory streams into compact latent codes that anticipate future observations. We let separate agents acquire such

← Previous
1…334335336337338…372
Next →