AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

PAT3D: Physics-Augmented Text-to-3D Scene Generation

DGX agent

arXiv:2511.21978v2 Announce Type: replace Abstract: We introduce PAT3D, the first physics-augmented text-to-3D scene generation framework that integrates vision-language models with physics-based simu

researcharxiv-cs-cv
24 Apr 2026
Research

Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2604.21536v1 Announce Type: cross Abstract: Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantic

researcharxiv-cs-ai
24 Apr 2026
Research

Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

DGX agent

arXiv:2604.21481v1 Announce Type: new Abstract: Crowdsourced pairwise evaluation has emerged as a scalable approach for assessing foundation models. However, applying it to Text to Speech(TTS) introdu

researcharxiv-cs-cl
24 Apr 2026
Research

Propensity Inference: Environmental Contributors to LLM Behaviour

DGX agent

arXiv:2604.21098v1 Announce Type: new Abstract: Motivated by loss of control risks from misaligned AI systems, we develop and apply methods for measuring language models' propensity for unsanctioned b

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning

DGX agent

arXiv:2604.21249v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) enable high-level semantic reasoning for end-to-end autonomous driving, particularly in unstructured environments, e

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

DGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Remember o3 was only a year and a week ago! Also, only GPT-5.5 seemed to take the 'evolution' piece seriously and change the setting rather …

DGX agent

I cannot provide an accurate summary for this entry as the text appears incomplete and lacks sufficient context. The post fragment references o3 (likely an AI model), GPT-5.5, and discusses timeline/e

model-releasesethan-mollick--x
24 Apr 2026
Research

Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs

DGX agent

arXiv:2604.20937v1 Announce Type: new Abstract: Video Large Language Models (Video LLMs) incur high inference latency due to a large number of visual tokens provided to LLMs. To address this, training

researcharxiv-cs-lg
24 Apr 2026
Model Releases

SparseGF: A Height-Aware Sparse Segmentation Framework with Context Compression for Robust Ground Filtering Across Urban to Natural Scenes

DGX agent

arXiv:2604.21356v1 Announce Type: new Abstract: High-quality digital terrain models derived from airborne laser scanning (ALS) data are essential for a wide range of geospatial analyses, and their gen

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection

DGX agent

arXiv:2604.21282v1 Announce Type: cross Abstract: Automated code vulnerability detection is critical for software security, yet existing approaches face a fundamental trade-off between detection accur

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

The Coding Assistant Breakdown: More Tokens Please

DGX agent

This analysis examines the token consumption and economics of coding assistants, likely comparing different AI models' efficiency and cost-effectiveness for code generation tasks. The piece probably d

model-releasessemianalysis
24 Apr 2026
Model Releases

The Root Theorem of Context Engineering

DGX agent

arXiv:2604.20874v1 Announce Type: cross Abstract: Every system that maintains a large language model conversation beyond a single session faces two inescapable constraints: the context window is finit

model-releasesarxiv-cs-cl
24 Apr 2026
Research

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping

DGX agent

arXiv:2604.21057v1 Announce Type: new Abstract: The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling L

researcharxiv-cs-cl
24 Apr 2026
Research

Transferable SCF-Acceleration through Solver-Aligned Initialization Learning

DGX agent

arXiv:2604.21657v1 Announce Type: new Abstract: Machine learning methods that predict initial guesses from molecular geometry can reduce this cost, but matrix-prediction models fail when extrapolating

researcharxiv-cs-lg
24 Apr 2026
Research

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping

DGX agent

arXiv:2511.13587v2 Announce Type: replace-cross Abstract: Visual autoregressive (AR) generation models have demonstrated strong potential for image generation, yet their next-token-prediction paradigm

researcharxiv-cs-ai
24 Apr 2026
Model Releases

We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through ParseBench, our comprehensive OCR benchmark over enterprise document…

DGX agent

We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through ParseBench, our comprehensive OCR benchmark over enterprise documents. We evaluated metrics across various dimensions: visual grou

model-releasesjerry-liu--x
24 Apr 2026
Model Releases

When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA

DGX agent

arXiv:2511.01458v2 Announce Type: replace-cross Abstract: Safety and reliability are critical for deploying visual question answering (VQA) systems in surgery, where incorrect or ambiguous responses c

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
Research

Analyzing Shapley Additive Explanations to Understand Anomaly Detection Algorithm Behaviors and Their Complementarity

DGX agent

arXiv:2602.00208v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection is a challenging problem due to the diversity of data distributions and the lack of labels. Ensemble methods ar

researcharxiv-cs-ai
23 Apr 2026
Model Releases

API pricing will be 5 per 1 million input tokens and 30 per 1 million output tokens, with a 1 million context window. (Remember, you will …

DGX agent

OpenAI's API pricing structure charges 5 per 1 million input tokens and 30 per 1 million output tokens, with support for a 1 million token context window. This pricing model reflects the higher cost o

model-releasessam-altman--x
23 Apr 2026
Model Releases

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

DGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

DGX agent

arXiv:2512.15146v3 Announce Type: replace Abstract: Test-time reinforcement learning mitigates the reliance on annotated data by using majority voting results as pseudo-labels, emerging as a complemen

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

DGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

DGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

DGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

DGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

safetyarxiv-cs-ai
23 Apr 2026
Applications

Continuous Semantic Caching for Low-Cost LLM Serving

DGX agent

arXiv:2604.20021v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly popular, caching responses so that they can be reused by users with semantically similar queries h

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

Differentiable Conformal Training for LLM Reasoning Factuality

DGX agent

arXiv:2604.20098v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently hallucinate, limiting their reliability in critical applications. Conformal Prediction (CP) addresses this by ca

model-releasesarxiv-cs-lg
23 Apr 2026
Safety

Epistemic Constitutionalism Or: how to avoid coherence bias

DGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

safetyarxiv-cs-ai
23 Apr 2026
Agents

EvoAgent: An Evolvable Agent Framework with Skill Learning and Multi-Agent Delegation

DGX agent

arXiv:2604.20133v1 Announce Type: new Abstract: This paper proposes EvoAgent - an evolvable large language model (LLM) agent framework that integrates structured skill learning with a hierarchical sub

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

EvoForest: A Novel Machine-Learning Paradigm via Open-Ended Evolution of Computational Graphs

DGX agent

arXiv:2604.19761v1 Announce Type: new Abstract: Modern machine learning is still largely organized around a single recipe: choose a parameterized model family and optimize its weights. Although highly

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Extract PDF text in your browser with LiteParse for the web

DGX agent

LlamaIndex have a most excellent open source project called LiteParse, which provides a Node.js CLI tool for extracting text from PDFs. I got a version of LiteParse working entirely in the browser, us

model-releasessimon-willison
23 Apr 2026
Research

First time fine-tuning, need a sanity check — 3B or 7B for multi-task reasoning? [D]

DGX agent

This Reddit discussion post addresses a beginner's question about choosing between 3B and 7B parameter models for fine-tuning on multi-task reasoning problems. The post likely contains advice from exp

researchr-machinelearning
23 Apr 2026
Agents

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

DGX agent

arXiv:2603.09046v2 Announce Type: replace-cross Abstract: Device-side Large Language Models (LLMs) have witnessed explosive growth, offering higher privacy and availability compared to cloud-side LLMs

agentsarxiv-cs-lg
23 Apr 2026
Industry

GPT 5.5 coming today.

DGX agent

Rumors suggest OpenAI may release GPT-5.5 on April 23, 2026 , following accidental leaks of the model in OpenAI's Codex platform that showed GPT-5.5 alongside other unreleased models . Internally code

industryr-chatgpt
23 Apr 2026
Model Releases

GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT (The Verge)

DGX agent

The Verge: GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT — The new model ‘excels’ at tasks

model-releasestechmeme
23 Apr 2026
Model Releases

GPT-5.5 on ARC-AGI (Verified) ARC-AGI-2: - Max: 85.0%, 1.87 - High: 83.3%, 1.45 - Med: 70.4%, 0.86 - Low: 33%, 0.35 GPT-5.5 is now state…

DGX agent

GPT-5.5 achieved state-of-the-art performance on the ARC-AGI-2 benchmark, with scores ranging from 85.0% on maximum difficulty tasks to 33% on low difficulty tasks. The model demonstrated consistent i

model-releasesfrancois-chollet--x
23 Apr 2026
Model Releases

How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues

DGX agent

arXiv:2604.19783v1 Announce Type: new Abstract: Which persuasion strategies, if any, are associated with donation compliance? Answering this requires fine-grained strategy labels across a full corpus

model-releasesarxiv-cs-cl
23 Apr 2026
Research

HumanScore: Benchmarking Human Motions in Generated Videos

DGX agent

arXiv:2604.20157v1 Announce Type: new Abstract: Recent advances in model architectures, compute, and data scale have driven rapid progress in video generation, producing increasingly realistic content

researcharxiv-cs-cv
23 Apr 2026
Safety

Hybrid Policy Distillation for LLMs

DGX agent

arXiv:2604.20244v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a powerful paradigm for compressing large language models (LLMs), whose effectiveness depends on intertwined choices of

safetyarxiv-cs-ai
23 Apr 2026
Safety

i-WiViG: Interpretable Window Vision GNN

DGX agent

arXiv:2503.08321v2 Announce Type: replace Abstract: Vision graph neural networks have emerged as a popular approach for modeling the global and spatial context for image recognition. However, a signif

safetyarxiv-cs-cv
23 Apr 2026
Model Releases

important (and very jakub-coded) jakub quote:

DGX agent

important (and very jakub-coded) jakub quote: OpenAI Unveils GPT-5.5. Company Says Expect a Faster Model Release Pace 👀 OpenAI: 'We see pretty significant improvements in the short term, extremely sig

model-releasessam-altman--x
23 Apr 2026
Model Releases

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

DGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

model-releasessimon-willison--x
23 Apr 2026
Model Releases

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

DGX agent

arXiv:2604.20545v1 Announce Type: new Abstract: In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: ben

model-releasesarxiv-cs-ai
23 Apr 2026
Research

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings

DGX agent

arXiv:2604.19902v1 Announce Type: cross Abstract: We present MMCORE, a unified framework designed for multimodal image generation and editing. MMCORE leverages a pre-trained Vision-Language Model (VLM

researcharxiv-cs-ai
23 Apr 2026
Model Releases

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

DGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Rabies diagnosis in low-data settings: A comparative study on the impact of data augmentation and transfer learning

DGX agent

arXiv:2604.19823v1 Announce Type: cross Abstract: Rabies remains a major public health concern across many African and Asian countries, where accurate diagnosis is critical for effective epidemiologic

researcharxiv-cs-ai
23 Apr 2026
Model Releases

RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering

DGX agent

arXiv:2604.20738v1 Announce Type: new Abstract: We introduce a benchmark dataset for question answering and translation in bilingual Latin and English settings, containing about 7,800 question-answer

model-releasesarxiv-cs-cl
23 Apr 2026
← Previous
1…551552553554555…1371
Next →