AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Analyzing Shapley Additive Explanations to Understand Anomaly Detection Algorithm Behaviors and Their Complementarity

DGX agent

arXiv:2602.00208v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection is a challenging problem due to the diversity of data distributions and the lack of labels. Ensemble methods ar

researcharxiv-cs-ai
23 Apr 2026
Model Releases

API pricing will be 5 per 1 million input tokens and 30 per 1 million output tokens, with a 1 million context window. (Remember, you will …

DGX agent

OpenAI's API pricing structure charges 5 per 1 million input tokens and 30 per 1 million output tokens, with support for a 1 million token context window. This pricing model reflects the higher cost o

model-releasessam-altman--x
23 Apr 2026
Model Releases

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

DGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

DGX agent

arXiv:2512.15146v3 Announce Type: replace Abstract: Test-time reinforcement learning mitigates the reliance on annotated data by using majority voting results as pseudo-labels, emerging as a complemen

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

DGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

DGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

DGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

DGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

safetyarxiv-cs-ai
23 Apr 2026
Applications

Continuous Semantic Caching for Low-Cost LLM Serving

DGX agent

arXiv:2604.20021v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly popular, caching responses so that they can be reused by users with semantically similar queries h

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

Differentiable Conformal Training for LLM Reasoning Factuality

DGX agent

arXiv:2604.20098v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently hallucinate, limiting their reliability in critical applications. Conformal Prediction (CP) addresses this by ca

model-releasesarxiv-cs-lg
23 Apr 2026
Safety

Epistemic Constitutionalism Or: how to avoid coherence bias

DGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

safetyarxiv-cs-ai
23 Apr 2026
Agents

EvoAgent: An Evolvable Agent Framework with Skill Learning and Multi-Agent Delegation

DGX agent

arXiv:2604.20133v1 Announce Type: new Abstract: This paper proposes EvoAgent - an evolvable large language model (LLM) agent framework that integrates structured skill learning with a hierarchical sub

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

EvoForest: A Novel Machine-Learning Paradigm via Open-Ended Evolution of Computational Graphs

DGX agent

arXiv:2604.19761v1 Announce Type: new Abstract: Modern machine learning is still largely organized around a single recipe: choose a parameterized model family and optimize its weights. Although highly

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Extract PDF text in your browser with LiteParse for the web

DGX agent

LlamaIndex have a most excellent open source project called LiteParse, which provides a Node.js CLI tool for extracting text from PDFs. I got a version of LiteParse working entirely in the browser, us

model-releasessimon-willison
23 Apr 2026
Research

First time fine-tuning, need a sanity check — 3B or 7B for multi-task reasoning? [D]

DGX agent

This Reddit discussion post addresses a beginner's question about choosing between 3B and 7B parameter models for fine-tuning on multi-task reasoning problems. The post likely contains advice from exp

researchr-machinelearning
23 Apr 2026
Agents

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

DGX agent

arXiv:2603.09046v2 Announce Type: replace-cross Abstract: Device-side Large Language Models (LLMs) have witnessed explosive growth, offering higher privacy and availability compared to cloud-side LLMs

agentsarxiv-cs-lg
23 Apr 2026
Industry

GPT 5.5 coming today.

DGX agent

Rumors suggest OpenAI may release GPT-5.5 on April 23, 2026 , following accidental leaks of the model in OpenAI's Codex platform that showed GPT-5.5 alongside other unreleased models . Internally code

industryr-chatgpt
23 Apr 2026
Model Releases

GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT (The Verge)

DGX agent

The Verge: GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT — The new model ‘excels’ at tasks

model-releasestechmeme
23 Apr 2026
Model Releases

GPT-5.5 on ARC-AGI (Verified) ARC-AGI-2: - Max: 85.0%, 1.87 - High: 83.3%, 1.45 - Med: 70.4%, 0.86 - Low: 33%, 0.35 GPT-5.5 is now state…

DGX agent

GPT-5.5 achieved state-of-the-art performance on the ARC-AGI-2 benchmark, with scores ranging from 85.0% on maximum difficulty tasks to 33% on low difficulty tasks. The model demonstrated consistent i

model-releasesfrancois-chollet--x
23 Apr 2026
Model Releases

How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues

DGX agent

arXiv:2604.19783v1 Announce Type: new Abstract: Which persuasion strategies, if any, are associated with donation compliance? Answering this requires fine-grained strategy labels across a full corpus

model-releasesarxiv-cs-cl
23 Apr 2026
Research

HumanScore: Benchmarking Human Motions in Generated Videos

DGX agent

arXiv:2604.20157v1 Announce Type: new Abstract: Recent advances in model architectures, compute, and data scale have driven rapid progress in video generation, producing increasingly realistic content

researcharxiv-cs-cv
23 Apr 2026
Safety

Hybrid Policy Distillation for LLMs

DGX agent

arXiv:2604.20244v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a powerful paradigm for compressing large language models (LLMs), whose effectiveness depends on intertwined choices of

safetyarxiv-cs-ai
23 Apr 2026
Safety

i-WiViG: Interpretable Window Vision GNN

DGX agent

arXiv:2503.08321v2 Announce Type: replace Abstract: Vision graph neural networks have emerged as a popular approach for modeling the global and spatial context for image recognition. However, a signif

safetyarxiv-cs-cv
23 Apr 2026
Model Releases

important (and very jakub-coded) jakub quote:

DGX agent

important (and very jakub-coded) jakub quote: OpenAI Unveils GPT-5.5. Company Says Expect a Faster Model Release Pace 👀 OpenAI: 'We see pretty significant improvements in the short term, extremely sig

model-releasessam-altman--x
23 Apr 2026
Model Releases

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

DGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

model-releasessimon-willison--x
23 Apr 2026
Model Releases

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

DGX agent

arXiv:2604.20545v1 Announce Type: new Abstract: In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: ben

model-releasesarxiv-cs-ai
23 Apr 2026
Research

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings

DGX agent

arXiv:2604.19902v1 Announce Type: cross Abstract: We present MMCORE, a unified framework designed for multimodal image generation and editing. MMCORE leverages a pre-trained Vision-Language Model (VLM

researcharxiv-cs-ai
23 Apr 2026
Model Releases

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

DGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Rabies diagnosis in low-data settings: A comparative study on the impact of data augmentation and transfer learning

DGX agent

arXiv:2604.19823v1 Announce Type: cross Abstract: Rabies remains a major public health concern across many African and Asian countries, where accurate diagnosis is critical for effective epidemiologic

researcharxiv-cs-ai
23 Apr 2026
Model Releases

RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering

DGX agent

arXiv:2604.20738v1 Announce Type: new Abstract: We introduce a benchmark dataset for question answering and translation in bilingual Latin and English settings, containing about 7,800 question-answer

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

RExBench: Can coding agents autonomously implement AI research extensions?

DGX agent

arXiv:2506.22598v3 Announce Type: replace Abstract: Agents based on Large Language Models (LLMs) have shown promise for performing sophisticated software engineering tasks autonomously. In addition, t

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Self-Describing Structured Data with Dual-Layer Guidance: A Lightweight Alternative to RAG for Precision Retrieval in Large-Scale LLM Knowledge Navigation

DGX agent

arXiv:2604.19777v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit a well-documented positional bias when processing long input contexts: information in the middle of a context win

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Separately, we’ve also heard reports of issues with Opus 4.7 in Claude Code. The team is working on those and we’ll share more as we roll ou…

DGX agent

Claude's Opus 4.7 model was experiencing issues when used within Claude Code, and Anthropic's team was actively investigating the problems with plans to provide updates as they continued the rollout.

model-releasesboris-cherny--x
23 Apr 2026
Agents

Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine

DGX agent

arXiv:2604.20022v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous diagnostic agents, yet they conflate two fundamentally different capabilities: natural-l

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

tldr: claude code changed some harness settings which degraded perf. these small harness tweaks can matter a lot! 1. default reasoning high …

DGX agent

tldr: claude code changed some harness settings which degraded perf. these small harness tweaks can matter a lot! 1. default reasoning high -> medium 2. bug that accidentally evicted thinking blocks o

model-releasesharrison-chase--x
23 Apr 2026
Model Releases

To Know is to Construct: Schema-Constrained Generation for Agent Memory

DGX agent

arXiv:2604.20117v1 Announce Type: new Abstract: Constructivist epistemology argues that knowledge is actively constructed rather than passively copied. Despite the generative nature of Large Language

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Top 10 uses for Codex at work

DGX agent

Codex is OpenAI's AI model designed to understand and generate code, helping developers automate programming tasks and improve productivity. This guide from OpenAI likely outlines ten practical workpl

model-releasesopenai
23 Apr 2026
Research

Treatment, evidence, imitation, and chat

DGX agent

arXiv:2506.23040v4 Announce Type: replace-cross Abstract: Large language models are thought to have the potential to aid in medical decision making. This work investigates the degree to which this mig

researcharxiv-cs-ai
23 Apr 2026
Applications

VAN-AD: Visual Masked Autoencoder with Normalizing Flow For Time Series Anomaly Detection

DGX agent

arXiv:2603.26842v2 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is essential for maintaining the reliability and security of IoT-enabled service systems. Existing method

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience

DGX agent

arXiv:2604.19756v1 Announce Type: cross Abstract: Large language model (LLM) agents often suffer from high reasoning overhead, excessive token consumption, unstable execution, and inability to reuse p

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

A Controlled Benchmark of Visual State-Space Backbones with Domain-Shift and Boundary Analysis for Remote-Sensing Segmentation

DGX agent

arXiv:2604.18721v1 Announce Type: cross Abstract: Visual state-space models (SSMs) are increasingly promoted as efficient alternatives to Vision Transformers, yet their practical advantages remain unc

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

Adaptive Prompt Elicitation for Text-to-Image Generation

DGX agent

arXiv:2602.04713v2 Announce Type: replace-cross Abstract: Aligning text-to-image generation with user intent remains challenging, as users frequently provide ambiguous inputs and struggle with model i

safetyarxiv-cs-ai
22 Apr 2026
Research

AI-Based Detection of Temporal Changes in MR-Linac Images Acquired During Routine Prostate Radiotherapy

DGX agent

arXiv:2602.04983v2 Announce Type: replace-cross Abstract: Purpose: To investigate whether an AI-based method can detect subtle inter-fraction changes in MR-Linac images acquired during radiotherapy an

researcharxiv-cs-ai
22 Apr 2026
Model Releases

🚨 Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 🤯 It’s an open-source pipel…

DGX agent

🚨 Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 🤯 It’s an open-source pipeline that replicates the exact daily loop of an ML researcher.

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language

DGX agent

arXiv:2604.19508v1 Announce Type: new Abstract: This paper introduces extit{Bangla Key2Text}, a large-scale dataset of 2.6 million Bangla keyword--text pairs designed for keyword-driven text generatio

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews

DGX agent

arXiv:2604.19502v1 Announce Type: new Abstract: The rapid adoption of Large Language Models (LLMs) has spurred interest in automated peer review; however, progress is currently stifled by benchmarks t

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Bridging Semantics and Geometry: A Decoupled LVLM-SAM Framework for Reasoning Segmentation in Optical Remote Sensing

DGX agent

arXiv:2512.19302v2 Announce Type: replace Abstract: Large Vision--Language Models (LVLMs) hold great promise for advancing optical remote sensing (RS) analysis, yet existing reasoning segmentation fra

safetyarxiv-cs-cv
22 Apr 2026
← Previous
1…562563564565566…1395
Next →