AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
24 Apr 2026

Empirical Comparison of Agent Communication Protocols for Task Orchestration

Model ReleasesDGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval

Model ReleasesDGX agent

arXiv:2604.21229v1 Announce Type: cross Abstract: Large language model assistants are increasingly expected to retain and reason over information accumulated across many sessions. We introduce Engrama

Enhancing Science Classroom Discourse Analysis through Joint Multi-Task Learning for Reasoning-Component Classification


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.21137v1 Announce Type: cross Abstract: Analyzing the reasoning patterns of students in science classrooms is critical for understanding knowledge construction mechanism and improving instru

Evaluating AI Meeting Summaries with a Reusable Cross-Domain Pipeline

Model ReleasesDGX agent

arXiv:2604.21345v1 Announce Type: new Abstract: We present a reusable evaluation pipeline for generative AI applications, instantiated for AI meeting summaries and released with a public artifact pack

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

Model ReleasesDGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

Model ReleasesDGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding

Model ReleasesDGX agent

arXiv:2506.19579v3 Announce Type: replace-cross Abstract: Robotic scene understanding increasingly relies on Vision-Language Models (VLMs) to generate natural language descriptions of the environment.

Feature request for @huggingface - add a repository size option to the sort menu, I want to see the DeepSeek quantized models that take up t…

Model ReleasesDGX agent

Simon Willison requested that Hugging Face add a repository size sorting option to help users find and filter models by storage requirements, specifically mentioning interest in locating DeepSeek quan

Federated Co-tuning Framework for Large and Small Language Models

Model ReleasesDGX agent

arXiv:2411.11707v3 Announce Type: replace-cross Abstract: By adapting Large Language Models (LLMs) to domain-specific tasks or enriching them with domain-specific knowledge, we can fully harness the c

Federated Learning for Surgical Vision in Appendicitis Classification: Results of the FedSurg EndoVis 2024 Challenge

Model ReleasesDGX agent

arXiv:2510.04772v2 Announce Type: replace-cross Abstract: Developing generalizable surgical AI requires multi-institutional data, yet patient privacy constraints preclude direct data sharing, making F

Fine-Tuning Regimes Define Distinct Continual Learning Problems

Model ReleasesDGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc

Flow Matching for Conditional MRI-CT and CBCT-CT Image Synthesis

Model ReleasesDGX agent

arXiv:2510.04823v2 Announce Type: replace Abstract: Generating synthetic CT (sCT) from MRI or CBCT plays a crucial role in enabling MRI-only and CBCT-based adaptive radiotherapy, improving treatment p

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

Model ReleasesDGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

for the record, i put 'undecided' as the last choice, not the second choice, @nikitabier sorry i know this is minor but it should be a 2 sec…

Model ReleasesDGX agent

for the record, i put 'undecided' as the last choice, not the second choice, @nikitabier sorry i know this is minor but it should be a 2 second fix in cursor https://x.com/kr0der/status/20477027526678

From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media

Model ReleasesDGX agent

arXiv:2604.21786v1 Announce Type: new Abstract: Social media platforms have become primary arenas for climate communication, generating millions of images and posts that - if systematically analysed -

From Research Question to Scientific Workflow: Leveraging Agentic AI for Science Automation

Model ReleasesDGX agent

arXiv:2604.21910v1 Announce Type: new Abstract: Scientific workflow systems automate execution -- scheduling, fault tolerance, resource management -- but not the semantic translation that precedes it.

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure

Model ReleasesDGX agent

arXiv:2512.22274v2 Announce Type: replace Abstract: We introduce GeCo, a geometry-grounded metric for jointly detecting geometric deformation and occlusion-inconsistency artifacts in static scenes. By

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning

Model ReleasesDGX agent

arXiv:2604.21495v1 Announce Type: cross Abstract: Numerical reasoning over expert-domain tables often exhibits high in-domain accuracy but limited robustness to domain shift. Models trained with super

Geo-R1: Improving Few-Shot Geospatial Referring Expression Understanding with Reinforcement Fine-Tuning

Model ReleasesDGX agent

arXiv:2509.21976v3 Announce Type: replace-cross Abstract: Referring expression understanding in remote sensing poses unique challenges, as it requires reasoning over complex object-context relationshi

Geometric Characterisation and Structured Trajectory Surrogates for Clinical Dataset Condensation

Model ReleasesDGX agent

arXiv:2604.21638v1 Announce Type: new Abstract: Dataset condensation constructs compact synthetic datasets that retain the training utility of large real-world datasets, enabling efficient model devel

Geometric Monomial (GEM): a family of rational 2N-differentiable activation functions

Model ReleasesDGX agent

arXiv:2604.21677v1 Announce Type: cross Abstract: The choice of activation function plays a crucial role in the optimization and performance of deep neural networks. While the Rectified Linear Unit (R

GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation

Model ReleasesDGX agent

arXiv:2604.21501v1 Announce Type: new Abstract: Lithology classification in well logs is a fundamental geoscience data mining task that aims to infer rock types from multi dimensional geophysical sequ

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

Model ReleasesDGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

Model ReleasesDGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

Model ReleasesDGX agent

arXiv:2604.21901v1 Announce Type: cross Abstract: As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopt

GPT-5.5 and GPT-5.5 Pro are now available in Hermes Agent through the Nous Portal and OpenRouter providers! (alongside the direct openai oau…

Model ReleasesDGX agent

GPT-5.5 and GPT-5.5 Pro are now available in Hermes Agent through the Nous Portal and OpenRouter providers! (alongside the direct openai oauth provider from yesterday) GPT-5.5 and GPT-5.5 Pro are now

GPT-5.5 is a giant leap forward for handling ambiguity compared to previous GPT models. As Windsurf 2.0 focuses more on parallel agents, thi…

Model ReleasesDGX agent

GPT-5.5 is a giant leap forward for handling ambiguity compared to previous GPT models. As Windsurf 2.0 focuses more on parallel agents, this model is key for long-horizon tasks — it excels at underst

GPT-5.5 is now available in Cursor! It's currently the top model on CursorBench at 72.8%. We've partnered with OpenAI to offer it for 50% of…

Model ReleasesDGX agent

Cursor has integrated OpenAI's GPT-5.5 model into its IDE, positioning it as the top performer on CursorBench with a 72.8% score. Through a partnership with OpenAI, the model is being offered at a 50%

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more aut…

Model ReleasesDGX agent

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than any GPT model we've tested, surfacing bugs no

GPT-5.5 is now available in Windsurf 2.0!

Model ReleasesDGX agent

Windsurf 2.0 has been released with integration of GPT-5.5, representing an upgrade to the Windsurf development platform's AI capabilities. This update likely provides users access to an advanced lang

GPT-5.5 is now available on Perplexity for Max subscribers. GPT-5.5 is also rolling out as the default orchestration model in Computer for b…

Model ReleasesDGX agent

Perplexity has made GPT-5.5 available to Max subscribers and is rolling it out as the default orchestration model in Perplexity Computer. The announcement indicates expanded access to OpenAI's GPT-5.5

GPT-5.5 now available in Deep Agents!

Model ReleasesDGX agent

GPT-5.5 now available in Deep Agents! GPT-5.5 is now available in the API. The model brings higher intelligence and stronger token efficiency to complex work, helping tasks get done with fewer retries

gpt-5.5 unlocks a new level of possibility:

Model ReleasesDGX agent

gpt-5.5 unlocks a new level of possibility: GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than a

Great launches recently. Time for temperature check! Based on what you've read, which of these 2 are you going to be using as your 'main' co…

Model ReleasesDGX agent

The content highlights technical issues encountered when attempting to access x.com. Users are warned that JavaScript must be enabled or a supported browser must be used, as the current settings are b

Grok is better than Gemini even in Hebrew. I asked exactly the same question about the Hubble Space Telescope, which was launched into space…

Model ReleasesDGX agent

Grok is better than Gemini even in Hebrew. I asked exactly the same question about the Hubble Space Telescope, which was launched into space 36 years ago today. Grok gave me a great summary with image

Grounding Machine Creativity in Game Design Knowledge Representations: Empirical Probing of LLM-Based Executable Synthesis of Goal Playable Patterns under Structural Constraints

Model ReleasesDGX agent

arXiv:2603.07101v3 Announce Type: replace Abstract: Creatively translating complex gameplay ideas into executable artifacts (e.g., games as Unity projects and code) remains a central challenge in comp

Grounding Video Reasoning in Physical Signals

Model ReleasesDGX agent

arXiv:2604.21873v1 Announce Type: new Abstract: Physical video understanding requires more than naming an event correctly. A model can answer a question about pouring, sliding, or collision from textu

Here's DeepSeek v4 Pro. Added to the playable gallery as well.

Model ReleasesDGX agent

Here's DeepSeek v4 Pro. Added to the playable gallery as well. Media I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 30

Huawei says its Ascend supernode based on the Ascend 950 AI chips will fully support DeepSeek V4, as DeepSeek launches a preview of its V4 model (Reuters)

Model ReleasesDGX agent

Reuters: Huawei says its Ascend supernode based on the Ascend 950 AI chips will fully support DeepSeek V4, as DeepSeek launches a preview of its V4 model — Huawei Technologies said on Friday its Ascen

huggingface: https://huggingface.co/collections/deepseek-ai/deepseek-v4.

Model ReleasesDGX agent

DeepSeek-V4 is a collection of models released by DeepSeek-AI on Hugging Face that represents their latest generation of large language models. The collection likely includes various model sizes and c

@huggingface On this page: https://huggingface.co/models?other=base_model:quantized:deepseek-ai/DeepSeek-V4-Flash

Model ReleasesDGX agent

This post references a Hugging Face Models page filtered to show quantized versions of the DeepSeek-V4-Flash model, a lightweight variant of DeepSeek's V4 language model. The page displays community-q

@hwchase17 Is there any deepagents skill similar to the claude-api skill to help me build apps with deepagents better? https://github.com/an…

Model ReleasesDGX agent

A user asks Harrison Chase about whether there are deepagents skills comparable to the claude-api skill that could help with building applications using deepagents framework. This appears to be a ques

HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks

Model ReleasesDGX agent

arXiv:2604.14709v2 Announce Type: replace Abstract: Existing benchmarks for hardware design primarily evaluate Large Language Models (LLMs) on isolated, component-level tasks such as generating HDL mo

HyperAdapt: Simple High-Rank Adaptation

Model ReleasesDGX agent

arXiv:2509.18629v3 Announce Type: replace-cross Abstract: Foundation models excel across diverse tasks, but adapting them to specialized applications often requires fine-tuning, an approach that is me

HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping

Model ReleasesDGX agent

arXiv:2604.21127v1 Announce Type: new Abstract: The NASA PACE mission provides unprecedented hyperspectral observations of ocean color, aerosols, and clouds, offering new insights into how these compo

Hyperloop Transformers

Model ReleasesDGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 3000 AD' in…

Model ReleasesDGX agent

I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 3000 AD' in one prompt. You can play the full gallery here: https://hg-

I hope the upgrade to DeepSeek v4 will make the bot comments on here more bearable.

Model ReleasesDGX agent

This post expresses hope that upgrading to DeepSeek v4 (an AI model) will improve the quality of bot-generated comments on a platform or service. The statement implies current bot comments are conside

ICNN-enhanced 2SP: Leveraging input convex neural networks for solving two-stage stochastic programming

Model ReleasesDGX agent

arXiv:2505.05261v3 Announce Type: replace-cross Abstract: Two-stage stochastic programming (2SP) offers a basic framework for modelling decision-making under uncertainty, yet scalability remains a cha

Ideological Bias in LLMs' Economic Causal Reasoning

Model ReleasesDGX agent

arXiv:2604.21334v1 Announce Type: new Abstract: Do large language models (LLMs) exhibit systematic ideological bias when reasoning about economic causal effects? As LLMs are increasingly used in polic

ILDR: Geometric Early Detection of Grokking

Model ReleasesDGX agent

arXiv:2604.20923v1 Announce Type: new Abstract: Grokking describes a delayed generalization phenomenon in which a neural network achieves perfect training accuracy long before validation accuracy impr

Instacart co-founder Apoorva Mehta launches Abundance, a hedge fund that aims to have AI agents run the entire fund, with $100M in seed funding (Hema Parmar/Bloomberg)

Model ReleasesDGX agent

Hema Parmar / Bloomberg: Instacart co-founder Apoorva Mehta launches Abundance, a hedge fund that aims to have AI agents run the entire fund, with $100M in seed funding — Instacart co-founder Apoorva

Intent Laundering: AI Safety Datasets Are Not What They Seem

Model ReleasesDGX agent

arXiv:2602.16729v3 Announce Type: replace-cross Abstract: We systematically evaluate the quality of widely used adversarial safety datasets from two perspectives: in isolation and in practice. In isol

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

Model ReleasesDGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can …

Model ReleasesDGX agent

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit from reli

IRIS: Interpolative Renyi Iterative Self-play for Large Language Model Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.20933v1 Announce Type: cross Abstract: Self-play fine-tuning enables large language models to improve beyond supervised fine-tuning without additional human annotations by contrasting annot

Is anyone using models to describe an image and get a prompt? Is there much difference between Qwen 3.5 9b vs Qwen 3.5 27b, vs gemma 4 27b and another model you use ?

Model ReleasesDGX agent

I'd need to search for this specific Reddit discussion to provide an accurate summary of what was actually discussed. Let me retrieve that information. This Reddit post discusses using AI vision model

It's High Time: A Survey of Temporal Question Answering

Model ReleasesDGX agent

arXiv:2505.20243v4 Announce Type: replace Abstract: Time plays a critical role in how information is generated, retrieved, and interpreted. In this survey, we provide a comprehensive overview of Tempo

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

Model ReleasesDGX agent

arXiv:2603.28342v2 Announce Type: replace Abstract: We present Kernel-Smith, a framework for high-performance GPU kernel and operator generation that combines a stable evaluation-driven evolutionary a

KompeteAI: Accelerated Autonomous Multi-Agent System for End-to-End Pipeline Generation for Machine Learning Problems

Model ReleasesDGX agent

arXiv:2508.10177v3 Announce Type: replace Abstract: Recent Large Language Model (LLM)-based AutoML systems demonstrate impressive capabilities but face significant limitations such as constrained expl

← Previous
1…311312313314315…373
Next →