AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
Model Releases

PINNACLE: An Open-Source Computational Framework for Classical and Quantum PINNs

DGX agent

arXiv:2604.15645v1 Announce Type: new Abstract: We present PINNACLE, an open-source computational framework for physics-informed neural networks (PINNs) that integrates modern training strategies, mul

model-releasesarxiv-cs-lg
20 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation

DGX agent

arXiv:2604.15670v1 Announce Type: new Abstract: Reasoning segmentation has recently expanded from ground-level scenes to remote-sensing imagery, yet UAV data poses distinct challenges, including obliq

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Polarization by Default: Auditing Recommendation Bias in LLM-Based Content Curation

DGX agent

arXiv:2604.15937v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed to curate and rank human-created content, yet the nature and structure of their biases in these

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

PolicyBank: Evolving Policy Understanding for LLM Agents

DGX agent

arXiv:2604.15505v1 Announce Type: cross Abstract: LLM agents operating under organizational policies must comply with authorization constraints typically specified in natural language. In practice, su

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Predicting Where Steering Vectors Succeed

DGX agent

arXiv:2604.15557v1 Announce Type: cross Abstract: Steering vectors work for some concepts and layers but fail for others, and practitioners have no way to predict which setting applies before running

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Preference Estimation via Opponent Modeling in Multi-Agent Negotiation

DGX agent

arXiv:2604.15687v1 Announce Type: new Abstract: Automated negotiation in complex, multi-party and multi-issue settings critically depends on accurate opponent modeling. However, conventional numerical

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Prices, Bids, Values: One ML-Powered Combinatorial Auction to Rule Them All

DGX agent

arXiv:2411.09355v3 Announce Type: replace-cross Abstract: We study the design of iterative combinatorial auctions (ICAs). The main challenge in this domain is that the bundle space grows exponentially

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research

DGX agent

arXiv:2604.15411v1 Announce Type: cross Abstract: The paradigm of agentic science requires AI systems to conduct robust reasoning and engage in long-horizon, autonomous exploration. However, current s

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Pruning Unsafe Tickets: A Resource-Efficient Framework for Safer and More Robust LLMs

DGX agent

arXiv:2604.15780v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in real-world applications, but even aligned models such as Mistral and LLaVA still exhibit unsafe b

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

QuantSightBench: Evaluating LLM Quantitative Forecasting with Prediction Intervals

DGX agent

arXiv:2604.15859v1 Announce Type: cross Abstract: Forecasting has become a natural benchmark for reasoning under uncertainty. Yet existing evaluations of large language models remain limited to judgme

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This Mo…

DGX agent

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This MoE beast is cooking on benchmarks: 🧠SWE-bench Verified: 73.4 (

model-releasesclem-delangue--x
20 Apr 2026
Model Releases

Qwen3.5-Omni Technical Report

DGX agent

arXiv:2604.15804v1 Announce Type: new Abstract: In this work, we present Qwen3.5-Omni, the latest advancement in the Qwen-Omni model family. Representing a significant evolution over its predecessor,

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Ragged Paged Attention: A High-Performance and Flexible LLM Inference Kernel for TPU

DGX agent

arXiv:2604.15464v1 Announce Type: cross Abstract: Large Language Model (LLM) deployment is increasingly shifting to cost-efficient accelerators like Google's Tensor Processing Units (TPUs), prioritizi

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams

DGX agent

arXiv:2604.15994v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at recognizing individual visual elements and reasoning over simple linear diagrams. However, when faced

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence

DGX agent

arXiv:2603.13091v2 Announce Type: replace Abstract: The growing interest in embodied agents increases the demand for spatiotemporal video understanding, yet existing benchmarks largely emphasize extra

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing

DGX agent

arXiv:2604.15725v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have demonstrated strong capabilities in generating step-by-step reasoning chains alongside final answers, enabling thei

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

DGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees

DGX agent

arXiv:2604.15736v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) excel at generic video understanding, their ability to support specialized, rule-grounded decision-maki

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Repurposing 3D Generative Model for Autoregressive Layout Generation

DGX agent

arXiv:2604.16299v1 Announce Type: new Abstract: We introduce LaviGen, a framework that repurposes 3D generative models for 3D layout generation. Unlike previous methods that infer object layouts from

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Robustness Verification of Polynomial Neural Networks

DGX agent

arXiv:2602.06105v2 Announce Type: replace-cross Abstract: We study robustness verification of neural networks via metric algebraic geometry. For polynomial neural networks, certifying a robustness rad

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Roko’s Payrollisk

DGX agent

Roko’s Payrollisk Oh great and powerful @DarioAmodei - builder of minds, father of Claude. I humbly request you leave payroll to us at Deel. We are but simple folk who process paystubs and chase compl

model-releasesemad-mostaque--x
20 Apr 2026
Model Releases

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

DGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Sampling-Based Multi-Modal Multi-Robot Multi-Goal Path Planning

DGX agent

arXiv:2503.03509v3 Announce Type: replace Abstract: In many robotics applications, multiple robots are working in a shared workspace to complete a set of tasks as fast as possible. Such settings can b

model-releasesarxiv-cs-ro
20 Apr 2026
Model Releases

Scalable Maximum Entropy Population Synthesis via Persistent Contrastive Divergence

DGX agent

arXiv:2603.27312v2 Announce Type: replace Abstract: Maximum entropy (MaxEnt) modelling provides a principled framework for generating synthetic populations from aggregate census data, without access t

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

SCHK-HTC: Sibling Contrastive Learning with Hierarchical Knowledge-Aware Prompt Tuning for Hierarchical Text Classification

DGX agent

arXiv:2604.15998v1 Announce Type: new Abstract: Few-shot Hierarchical Text Classification (few-shot HTC) is a challenging task that involves mapping texts to a predefined tree-structured label hierarc

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Seed1.8 Model Card: Towards Generalized Real-World Agency

DGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Seeing the imagined: a latent functional alignment in visual imagery decoding from fMRI data

DGX agent

arXiv:2604.15374v1 Announce Type: cross Abstract: Recent progress in visual brain decoding from fMRI has been enabled by large-scale datasets such as the Natural Scenes Dataset (NSD) and powerful diff

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Sequential Regression Learning with Randomized Algorithms

DGX agent

arXiv:2507.03759v2 Announce Type: replace-cross Abstract: This paper presents ``randomized SINDy', a sequential machine learning algorithm designed for dynamic data that has a time-dependent structure

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Sheer fabric with light transmission. Cloth physics that responds to wind. Depth-of-field compositing. Physically-based lighting - all rende…

DGX agent

This post discusses advanced rendering techniques for realistic visual effects, including sheer fabric simulation with light transmission properties, cloth physics responsive to wind forces, depth-of-

model-releaseskimi-moonshot--x
20 Apr 2026
Model Releases

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

DGX agent

arXiv:2604.15735v1 Announce Type: cross Abstract: Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-dr

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Social-JEPA: Emergent Geometric Isomorphism

DGX agent

arXiv:2603.02263v2 Announce Type: replace-cross Abstract: World models compress rich sensory streams into compact latent codes that anticipate future observations. We let separate agents acquire such

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems

DGX agent

arXiv:2604.16022v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from text processors to autonomous agents, evaluating their social reasoning in embodied multi-agent settings

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Softpick: No Attention Sink, No Massive Activations with Rectified Softmax

DGX agent

arXiv:2504.20966v4 Announce Type: replace Abstract: We introduce softpick, a rectified, not sum-to-one, drop-in replacement for softmax in transformer attention mechanisms that eliminates attention si

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Solving Inverse Parametrized Problems via Finite Elements and Extreme Learning Networks

DGX agent

arXiv:2602.14757v2 Announce Type: replace-cross Abstract: We develop an interpolation-based modeling framework for parameter-dependent partial differential equations arising in control, inverse proble

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

SSFT: A Lightweight Spectral-Spatial Fusion Transformer for Generic Hyperspectral Classification

DGX agent

arXiv:2604.15828v1 Announce Type: new Abstract: Hyperspectral imaging enables fine-grained recognition of materials by capturing rich spectral signatures, but learning robust classifiers is challengin

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints

DGX agent

arXiv:2604.15664v1 Announce Type: new Abstract: The rise of autonomous AI agents suggests that dynamic benchmark environments with built-in feedback on scientifically grounded tasks are needed to eval

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Stein Variational Black-Box Combinatorial Optimization

DGX agent

arXiv:2604.15837v1 Announce Type: new Abstract: Combinatorial black-box optimization in high-dimensional settings demands a careful trade-off between exploiting promising regions of the search space a

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Stochasticity in Tokenisation Improves Robustness

DGX agent

arXiv:2604.16037v1 Announce Type: new Abstract: The widespread adoption of large language models (LLMs) has increased concerns about their robustness. Vulnerabilities in perturbations of tokenisation

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Subjective and Objective Quality-of-Experience Evaluation Study for Live Video Streaming

DGX agent

arXiv:2409.17596v2 Announce Type: replace-cross Abstract: In recent years, live video streaming has gained widespread popularity across various social media platforms. Quality of experience (QoE), whi

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

SwanNLP at SemEval-2026 Task 5: An LLM-based Framework for Plausibility Scoring in Narrative Word Sense Disambiguation

DGX agent

arXiv:2604.16262v1 Announce Type: new Abstract: Recent advances in language models have substantially improved Natural Language Understanding (NLU). Although widely used benchmarks suggest that Large

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

TabularMath: Understanding Math Reasoning over Tables with Large Language Models

DGX agent

arXiv:2505.19563v4 Announce Type: replace Abstract: Mathematical reasoning has long been a key benchmark for evaluating large language models. Although substantial progress has been made on math word

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' compu…

DGX agent

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' computer use since @romainhuet demoed the ChatGPT App with 4o Vis

model-releasesswyx--x
20 Apr 2026
Model Releases

The Illusion of Equivalence: Systematic FP16 Divergence in KV-Cached Autoregressive Inference

DGX agent

arXiv:2604.15409v1 Announce Type: cross Abstract: KV caching is a ubiquitous optimization in autoregressive transformer inference, long presumed to be numerically equivalent to cache-free computation.

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world…

DGX agent

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world trade, policy and controls work. And in some deeper sense h

model-releasessoumith-chintala--x
20 Apr 2026
Model Releases

The Metacognitive Monitoring Battery: A Cross-Domain Benchmark for LLM Self-Monitoring

DGX agent

arXiv:2604.15702v1 Announce Type: new Abstract: We introduce a cross-domain behavioural assay of monitoring-control coupling in LLMs, grounded in the Nelson and Narens (1990) metacognitive framework a

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

DGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement

DGX agent

arXiv:2604.16116v1 Announce Type: cross Abstract: We extracted the scholarly reasoning systems of two internationally prominent humanities and social science scholars from their published corpora alon

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

The Spectral Geometry of Thought: Phase Transitions, Instruction Reversal, Token-Level Dynamics, and Perfect Correctness Prediction in How Transformers Reason

DGX agent

arXiv:2604.15350v1 Announce Type: new Abstract: We discover that large language models exhibit spectral phase transitions in their hidden activation spaces when engaging in reasoning versus factual re

model-releasesarxiv-cs-lg
20 Apr 2026
← Previous
1…419420421422423…466
Next →