AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
Model Releases

Dual Causal Inference: Integrating Backdoor Adjustment and Instrumental Variable Learning for Medical VQA

DGX agent

arXiv:2604.20306v1 Announce Type: cross Abstract: Medical Visual Question Answering (MedVQA) aims to generate clinically reliable answers conditioned on complex medical images and questions. However,

model-releasesarxiv-cs-ai
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Duluth at SemEval-2026 Task 6: DeBERTa with LLM-Augmented Data for Unmasking Political Question Evasions

DGX agent

arXiv:2604.20168v1 Announce Type: new Abstract: This paper presents the Duluth approach to SemEval-2026 Task 6 on CLARITY: Unmasking Political Question Evasions. We address Task 1 (clarity-level class

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis

DGX agent

arXiv:2604.20523v1 Announce Type: cross Abstract: We study whether Large Language Models (LLMs) can perform feature model analysis operations (AOs) directly on semi-formal textual blueprints, i.e., co

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Earth Day at #GoogleCloudNext, I’m demoing a Sustainability Agent at the @nvidia booth. Built with @Google ADK, @googlegemma, #nemotron, Clo…

DGX agent

Earth Day at #GoogleCloudNext, I’m demoing a Sustainability Agent at the @nvidia booth. Built with @Google ADK, @googlegemma, #nemotron, Cloud Run, @milvusio , @LangChain & @ollama to reason across im

model-releasesollama--x
23 Apr 2026
Model Releases

Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

DGX agent

arXiv:2511.06209v4 Announce Type: replace Abstract: LLMs can solve complex tasks by generating long, multi-step reasoning chains. Test-time scaling (TTS) can further improve LLM performance by samplin

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

embers

DGX agent

embers GPT-5.5, not fully saturating the TikZ unicorn test yet but getting awfully close ... (yes this is actual TikZ code, I personally find it so unbelievable that I'm putting the code below for any

model-releasessam-altman--x
23 Apr 2026
Model Releases

Evaluating the Quality of the Quantified Uncertainty for (Re)Calibration of Data-Driven Regression Models

DGX agent

arXiv:2508.17761v3 Announce Type: replace Abstract: In safety-critical applications data-driven models must not only be accurate but also provide reliable uncertainty estimates. This property, commonl

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Evian: Towards Explainable Visual Instruction-tuning Data Auditing

DGX agent

arXiv:2604.20544v1 Announce Type: cross Abstract: The efficacy of Large Vision-Language Models (LVLMs) is critically dependent on the quality of their training data, requiring a precise balance betwee

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Evidence of Layered Positional and Directional Constraints in the Voynich Manuscript: Implications for Cipher-Like Structure

DGX agent

arXiv:2604.19762v1 Announce Type: new Abstract: The Voynich Manuscript (VMS) exhibits a script of uncertain origin whose grapheme sequences have resisted linguistic analysis. We present a systematic a

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

EvoForest: A Novel Machine-Learning Paradigm via Open-Ended Evolution of Computational Graphs

DGX agent

arXiv:2604.19761v1 Announce Type: new Abstract: Modern machine learning is still largely organized around a single recipe: choose a parameterized model family and optimize its weights. Although highly

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts

DGX agent

arXiv:2604.19835v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for scaling large language models: frontier models routinely decouple total parameters f

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization

DGX agent

arXiv:2604.20726v1 Announce Type: cross Abstract: This work explores the role of prompt design and judge selection in LLM-as-a-Judge evaluations of free text legal question answering. We examine wheth

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Exploring Data Augmentation and Resampling Strategies for Transformer-Based Models to Address Class Imbalance in AI Scoring of Scientific Explanations in NGSS Classroom

DGX agent

arXiv:2604.19754v1 Announce Type: new Abstract: Automated scoring of students' scientific explanations offers the potential for immediate, accurate feedback, yet class imbalance in rubric categories p

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Exploring Spatial Intelligence from a Generative Perspective

DGX agent

arXiv:2604.20570v1 Announce Type: new Abstract: Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective.

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Extract PDF text in your browser with LiteParse for the web

DGX agent

LlamaIndex have a most excellent open source project called LiteParse, which provides a Node.js CLI tool for extracting text from PDFs. I got a version of LiteParse working entirely in the browser, us

model-releasessimon-willison
23 Apr 2026
Model Releases

Fairness Testing of Large Language Models in Role-Playing

DGX agent

arXiv:2411.00585v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become foundational in modern language-driven software applications, profoundly influencing daily life. A cr

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Falcon 9 launches 24 @Starlink satellites from California

DGX agent

SpaceX's Falcon 9 rocket successfully launched 24 Starlink satellites from a California launch facility, continuing the company's ongoing deployment of its satellite internet constellation. This missi

model-releaseselon-musk--x
23 Apr 2026
Model Releases

Fast Bayesian equipment condition monitoring via simulation based inference: applications to heat exchanger health

DGX agent

arXiv:2604.20735v1 Announce Type: new Abstract: Accurate condition monitoring of industrial equipment requires inferring latent degradation parameters from indirect sensor measurements under uncertain

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing

DGX agent

arXiv:2604.20429v1 Announce Type: new Abstract: Remote sensing (RS) image-text retrieval plays a critical role in understanding massive RS imagery. However, the dense multi-object distribution and com

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

FeDa4Fair: Client-Level Federated Datasets for Fairness Evaluation

DGX agent

arXiv:2506.21095v4 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative training while preserving privacy, yet it introduces a critical challenge: the 'illusion of fair

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Fextsuperscript{2}LP-AP: Fast & Flexible Label Propagation with Adaptive Propagation Kernel

DGX agent

arXiv:2604.20736v1 Announce Type: new Abstract: Semi-supervised node classification is a foundational task in graph machine learning, yet state-of-the-art Graph Neural Networks (GNNs) are hindered by

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Finding Duplicates in 1.1M BDD Steps: cukereuse, a Paraphrase-Robust Static Detector for Cucumber and Gherkin

DGX agent

arXiv:2604.20462v1 Announce Type: cross Abstract: Behaviour-Driven Development (BDD) suites accumulate step-text duplication whose maintenance cost is established in prior work. Existing detection tec

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

FlashNorm: Fast Normalization for Transformers

DGX agent

arXiv:2407.09577v4 Announce Type: replace Abstract: Normalization layers are ubiquitous in large language models (LLMs) yet represent a compute bottleneck: on hardware with distinct vector and matrix

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Foundation Models in Biomedical Imaging: Turning Hype into Reality

DGX agent

arXiv:2512.15808v2 Announce Type: replace-cross Abstract: Foundation models (FMs) are driving a prominent shift in biomedical imaging from task-specific models to unified backbone models for diverse t

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

DGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents

DGX agent

arXiv:2604.20006v1 Announce Type: new Abstract: Personalized agents that interact with users over long periods must maintain persistent memory across sessions and update it as circumstances change. Ho

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

From Scene to Object: Text-Guided Dual-Gaze Prediction

DGX agent

arXiv:2604.20191v1 Announce Type: cross Abstract: Interpretable driver attention prediction is crucial for human-like autonomous driving. However, existing datasets provide only scene-level global gaz

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Global Offshore Wind Infrastructure: Deployment and Operational Dynamics from Dense Sentinel-1 Time Series

DGX agent

arXiv:2604.20822v1 Announce Type: new Abstract: The offshore wind energy sector is expanding rapidly, increasing the need for independent, high-temporal-resolution monitoring of infrastructure deploym

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

GPT-5.5 Bio Bug Bounty

DGX agent

OpenAI's GPT-5.5 Bio Bug Bounty program invites security researchers to identify and report vulnerabilities in GPT-5.5's biological information handling capabilities, focusing on potential misuse risk

model-releasesopenai
23 Apr 2026
Model Releases

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality'…

DGX agent

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality' - Gets lots of stuff done without pausing unnecessarily It

model-releasesdair-ai--x
23 Apr 2026
Model Releases

GPT-5.5 is here! We hope it's useful to you. I personally like it.

DGX agent

I don't have verified information about a GPT-5.5 model release. This appears to be either a fictional or future-dated post, as it references a non-existent model and uses a URL format/status ID that

model-releasessam-altman--x
23 Apr 2026
Model Releases

GPT-5.5 is likely the best model in the world. But open models like Kimi and Minimax get almost identical coding benchmark scores at 10-25x …

DGX agent

GPT-5.5 is likely the best model in the world. But open models like Kimi and Minimax get almost identical coding benchmark scores at 10-25x lower cost. I broke down the benchmarks and pricing. Here's

model-releasesclem-delangue--x
23 Apr 2026
Model Releases

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get st…

DGX agent

GPT-5.5 is now accessible in Hermes Agent through the ChatGPT/Codex OAuth provider. Run `hermes update` to access now or learn how to get started with Hermes Agent here: https://hermes-agent.nousresea

model-releasesnous-research--x
23 Apr 2026
Model Releases

GPT-5.5 is priced at 5/1M input tokens and 30/1M output tokens, double GPT-5.4's pricing; GPT-5.5 Pro costs 30/1M input tokens and 180/1M output tokens (Carl Franzen/VentureBeat)

DGX agent

Carl Franzen / VentureBeat: GPT-5.5 is priced at 5/1M input tokens and 30/1M output tokens, double GPT-5.4's pricing; GPT-5.5 Pro costs 30/1M input tokens and 180/1M output tokens — After months of ru

model-releasestechmeme
23 Apr 2026
Model Releases

GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT (The Verge)

DGX agent

The Verge: GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT — The new model ‘excels’ at tasks

model-releasestechmeme
23 Apr 2026
Model Releases

GPT-5.5 may not be in the official OpenAI API... but it's available via the apparently approved-of Codex API backdoor So I used that to make…

DGX agent

GPT-5.5 may not be in the official OpenAI API... but it's available via the apparently approved-of Codex API backdoor So I used that to make these pelicans (default and xhigh)! https://simonwillison.n

model-releasessimon-willison--x
23 Apr 2026
Model Releases

GPT-5.5 on ARC-AGI (Verified) ARC-AGI-2: - Max: 85.0%, 1.87 - High: 83.3%, 1.45 - Med: 70.4%, 0.86 - Low: 33%, 0.35 GPT-5.5 is now state…

DGX agent

GPT-5.5 achieved state-of-the-art performance on the ARC-AGI-2 benchmark, with scores ranging from 85.0% on maximum difficulty tasks to 33% on low difficulty tasks. The model demonstrated consistent i

model-releasesfrancois-chollet--x
23 Apr 2026
Model Releases

Graph-Theoretic Models for the Prediction of Molecular Measurements

DGX agent

arXiv:2604.19840v1 Announce Type: new Abstract: Graph-theoretic approaches offer simplicity, interpretability, and low computational cost for molecular property prediction. Among these, the model prop

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Here’s my view on GPT-5.5, which I have been testing for a couple of weeks. It conducted not-bad social science research on its own, develop…

DGX agent

Here’s my view on GPT-5.5, which I have been testing for a couple of weeks. It conducted not-bad social science research on its own, developed a novel RPG & more. There is still jaggedness but GPT-5.5

model-releasesethan-mollick--x
23 Apr 2026
Model Releases

HiPO: Hierarchical Preference Optimization for Adaptive Reasoning in LLMs

DGX agent

arXiv:2604.20140v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is an effective framework for aligning large language models with human preferences, but it struggles with complex

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

How close are models to building a perfect Slack clone with less than $50,000 of tokens? I feel like not that far...

DGX agent

How close are models to building a perfect Slack clone with less than 50,000 of tokens? I feel like not that far... Claude Code spend had gotten to 10.95M runrate peak at SemiAnalysis But then Opus 4.

model-releasesdylan-patel--x
23 Apr 2026
Model Releases

How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues

DGX agent

arXiv:2604.19783v1 Announce Type: new Abstract: Which persuasion strategies, if any, are associated with donation compliance? Answering this requires fine-grained strategy labels across a full corpus

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Hybrid Multi-Phase Page Matching and Multi-Layer Diff Detection for Japanese Building Permit Document Review

DGX agent

arXiv:2604.19770v1 Announce Type: new Abstract: We present a hybrid multi-phase page matching algorithm for automated comparison of Japanese building permit document sets. Building permit review in Ja

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

I had early access to GPT-5.5. It is very good, especially the Pro version. Full writeup very shortly.

DGX agent

Ethan Mollick posted on X about having early access to GPT-5.5, commenting positively on its capabilities and noting that the Pro version is particularly strong. He indicated that a full detailed writ

model-releasesethan-mollick--x
23 Apr 2026
Model Releases

I’d been part of OpenAI early tester group for GPT-5.5. I believe with GPT-5.5 Pro we reached another inflection point-comparable to the ori…

DGX agent

I’d been part of OpenAI early tester group for GPT-5.5. I believe with GPT-5.5 Pro we reached another inflection point-comparable to the original release of o1-preview & then with 5.0 Pro, I had felt.

model-releasessam-altman--x
23 Apr 2026
Model Releases

If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most …

DGX agent

If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most comprehensive document OCR benchmark over real enterprise docu

model-releasesjerry-liu--x
23 Apr 2026
Model Releases

I'm a manager at @OpenAI, but with GPT-5.5 I'm a more effective IC than I've ever been. I can now write CUDA kernels like a pro. I can rely …

DGX agent

I'm a manager at @OpenAI, but with GPT-5.5 I'm a more effective IC than I've ever been. I can now write CUDA kernels like a pro. I can rely on it to run my research experiments. And we know how to mak

model-releasessam-altman--x
23 Apr 2026
Model Releases

IMPACT-CYCLE: A Contract-Based Multi-Agent System for Claim-Level Supervisory Correction of Long-Video Semantic Memory

DGX agent

arXiv:2604.20136v1 Announce Type: cross Abstract: Correcting errors in long-video understanding is disproportionately costly: existing multimodal pipelines produce opaque, end-to-end outputs that expo

model-releasesarxiv-cs-ai
23 Apr 2026
← Previous
1…399400401402403…469
Next →