AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
23 Apr 2026

EvoAgent: An Evolvable Agent Framework with Skill Learning and Multi-Agent Delegation

AgentsDGX agent

arXiv:2604.20133v1 Announce Type: new Abstract: This paper proposes EvoAgent - an evolvable large language model (LLM) agent framework that integrates structured skill learning with a hierarchical sub

EvoForest: A Novel Machine-Learning Paradigm via Open-Ended Evolution of Computational Graphs

Model ReleasesDGX agent

arXiv:2604.19761v1 Announce Type: new Abstract: Modern machine learning is still largely organized around a single recipe: choose a parameterized model family and optimize its weights. Although highly

Extract PDF text in your browser with LiteParse for the web

Model ReleasesDGX agent

LlamaIndex have a most excellent open source project called LiteParse, which provides a Node.js CLI tool for extracting text from PDFs. I got a version of LiteParse working entirely in the browser, us

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

First time fine-tuning, need a sanity check — 3B or 7B for multi-task reasoning? [D]

ResearchDGX agent

This Reddit discussion post addresses a beginner's question about choosing between 3B and 7B parameter models for fine-tuning on multi-task reasoning problems. The post likely contains advice from exp

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

AgentsDGX agent

arXiv:2603.09046v2 Announce Type: replace-cross Abstract: Device-side Large Language Models (LLMs) have witnessed explosive growth, offering higher privacy and availability compared to cloud-side LLMs

GPT 5.5 coming today.

IndustryDGX agent

Rumors suggest OpenAI may release GPT-5.5 on April 23, 2026 , following accidental leaks of the model in OpenAI's Codex platform that showed GPT-5.5 alongside other unreleased models . Internally code

GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT (The Verge)

Model ReleasesDGX agent

The Verge: GPT-5.5 is rolling out to Plus, Pro, Business, and Enterprise users in ChatGPT and Codex, and GPT-5.5 Pro to Pro, Business, and Enterprise users in ChatGPT — The new model ‘excels’ at tasks

GPT-5.5 on ARC-AGI (Verified) ARC-AGI-2: - Max: 85.0%, 1.87 - High: 83.3%, 1.45 - Med: 70.4%, 0.86 - Low: 33%, 0.35 GPT-5.5 is now state…

Model ReleasesDGX agent

GPT-5.5 achieved state-of-the-art performance on the ARC-AGI-2 benchmark, with scores ranging from 85.0% on maximum difficulty tasks to 33% on low difficulty tasks. The model demonstrated consistent i

How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues

Model ReleasesDGX agent

arXiv:2604.19783v1 Announce Type: new Abstract: Which persuasion strategies, if any, are associated with donation compliance? Answering this requires fine-grained strategy labels across a full corpus

HumanScore: Benchmarking Human Motions in Generated Videos

ResearchDGX agent

arXiv:2604.20157v1 Announce Type: new Abstract: Recent advances in model architectures, compute, and data scale have driven rapid progress in video generation, producing increasingly realistic content

Hybrid Policy Distillation for LLMs

SafetyDGX agent

arXiv:2604.20244v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a powerful paradigm for compressing large language models (LLMs), whose effectiveness depends on intertwined choices of

i-WiViG: Interpretable Window Vision GNN

SafetyDGX agent

arXiv:2503.08321v2 Announce Type: replace Abstract: Vision graph neural networks have emerged as a popular approach for modeling the global and spatial context for image recognition. However, a signif

important (and very jakub-coded) jakub quote:

Model ReleasesDGX agent

important (and very jakub-coded) jakub quote: OpenAI Unveils GPT-5.5. Company Says Expect a Faster Model Release Pace 👀 OpenAI: 'We see pretty significant improvements in the short term, extremely sig

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

Model ReleasesDGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

Model ReleasesDGX agent

arXiv:2604.20545v1 Announce Type: new Abstract: In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: ben

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings

ResearchDGX agent

arXiv:2604.19902v1 Announce Type: cross Abstract: We present MMCORE, a unified framework designed for multimodal image generation and editing. MMCORE leverages a pre-trained Vision-Language Model (VLM

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

Model ReleasesDGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

Rabies diagnosis in low-data settings: A comparative study on the impact of data augmentation and transfer learning

ResearchDGX agent

arXiv:2604.19823v1 Announce Type: cross Abstract: Rabies remains a major public health concern across many African and Asian countries, where accurate diagnosis is critical for effective epidemiologic

RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering

Model ReleasesDGX agent

arXiv:2604.20738v1 Announce Type: new Abstract: We introduce a benchmark dataset for question answering and translation in bilingual Latin and English settings, containing about 7,800 question-answer

RExBench: Can coding agents autonomously implement AI research extensions?

Model ReleasesDGX agent

arXiv:2506.22598v3 Announce Type: replace Abstract: Agents based on Large Language Models (LLMs) have shown promise for performing sophisticated software engineering tasks autonomously. In addition, t

Self-Describing Structured Data with Dual-Layer Guidance: A Lightweight Alternative to RAG for Precision Retrieval in Large-Scale LLM Knowledge Navigation

Model ReleasesDGX agent

arXiv:2604.19777v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit a well-documented positional bias when processing long input contexts: information in the middle of a context win

Separately, we’ve also heard reports of issues with Opus 4.7 in Claude Code. The team is working on those and we’ll share more as we roll ou…

Model ReleasesDGX agent

Claude's Opus 4.7 model was experiencing issues when used within Claude Code, and Anthropic's team was actively investigating the problems with plans to provide updates as they continued the rollout.

Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine

AgentsDGX agent

arXiv:2604.20022v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous diagnostic agents, yet they conflate two fundamentally different capabilities: natural-l

tldr: claude code changed some harness settings which degraded perf. these small harness tweaks can matter a lot! 1. default reasoning high …

Model ReleasesDGX agent

tldr: claude code changed some harness settings which degraded perf. these small harness tweaks can matter a lot! 1. default reasoning high -> medium 2. bug that accidentally evicted thinking blocks o

To Know is to Construct: Schema-Constrained Generation for Agent Memory

Model ReleasesDGX agent

arXiv:2604.20117v1 Announce Type: new Abstract: Constructivist epistemology argues that knowledge is actively constructed rather than passively copied. Despite the generative nature of Large Language

Top 10 uses for Codex at work

Model ReleasesDGX agent

Codex is OpenAI's AI model designed to understand and generate code, helping developers automate programming tasks and improve productivity. This guide from OpenAI likely outlines ten practical workpl

Treatment, evidence, imitation, and chat

ResearchDGX agent

arXiv:2506.23040v4 Announce Type: replace-cross Abstract: Large language models are thought to have the potential to aid in medical decision making. This work investigates the degree to which this mig

VAN-AD: Visual Masked Autoencoder with Normalizing Flow For Time Series Anomaly Detection

ApplicationsDGX agent

arXiv:2603.26842v2 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is essential for maintaining the reliability and security of IoT-enabled service systems. Existing method

WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience

Model ReleasesDGX agent

arXiv:2604.19756v1 Announce Type: cross Abstract: Large language model (LLM) agents often suffer from high reasoning overhead, excessive token consumption, unstable execution, and inability to reuse p

22 Apr 2026

A Controlled Benchmark of Visual State-Space Backbones with Domain-Shift and Boundary Analysis for Remote-Sensing Segmentation

Model ReleasesDGX agent

arXiv:2604.18721v1 Announce Type: cross Abstract: Visual state-space models (SSMs) are increasingly promoted as efficient alternatives to Vision Transformers, yet their practical advantages remain unc

Adaptive Prompt Elicitation for Text-to-Image Generation

SafetyDGX agent

arXiv:2602.04713v2 Announce Type: replace-cross Abstract: Aligning text-to-image generation with user intent remains challenging, as users frequently provide ambiguous inputs and struggle with model i

AI-Based Detection of Temporal Changes in MR-Linac Images Acquired During Routine Prostate Radiotherapy

ResearchDGX agent

arXiv:2602.04983v2 Announce Type: replace-cross Abstract: Purpose: To investigate whether an AI-based method can detect subtle inter-fraction changes in MR-Linac images acquired during radiotherapy an

🚨 Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 🤯 It’s an open-source pipel…

Model ReleasesDGX agent

🚨 Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 🤯 It’s an open-source pipeline that replicates the exact daily loop of an ML researcher.

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language

Model ReleasesDGX agent

arXiv:2604.19508v1 Announce Type: new Abstract: This paper introduces extit{Bangla Key2Text}, a large-scale dataset of 2.6 million Bangla keyword--text pairs designed for keyword-driven text generatio

Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews

Model ReleasesDGX agent

arXiv:2604.19502v1 Announce Type: new Abstract: The rapid adoption of Large Language Models (LLMs) has spurred interest in automated peer review; however, progress is currently stifled by benchmarks t

Bridging Semantics and Geometry: A Decoupled LVLM-SAM Framework for Reasoning Segmentation in Optical Remote Sensing

SafetyDGX agent

arXiv:2512.19302v2 Announce Type: replace Abstract: Large Vision--Language Models (LVLMs) hold great promise for advancing optical remote sensing (RS) analysis, yet existing reasoning segmentation fra

Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences

Model ReleasesDGX agent

arXiv:2601.04925v2 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive text, raising concerns about their misuse for propaganda, manipulation, and other harmfu

CoDA: Towards Effective Cross-domain Knowledge Transfer via CoT-guided Domain Adaptation

ApplicationsDGX agent

arXiv:2604.19488v1 Announce Type: new Abstract: Large language models (LLMs) have achieved substantial advances in logical reasoning, yet they continue to lag behind human-level performance. In-contex

CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers

SafetyDGX agent

arXiv:2604.19632v1 Announce Type: new Abstract: Graphic design images consist of multiple editable layers, such as text, background, and decorative elements, while most generative models produce raste

Dual Triangle Attention: Effective Bidirectional Attention Without Positional Embeddings

SafetyDGX agent

arXiv:2604.18603v1 Announce Type: cross Abstract: Bidirectional transformers are the foundation of many sequence modeling tasks across natural, biological, and chemical language domains, but they are

DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning

Model ReleasesDGX agent

arXiv:2604.18964v1 Announce Type: new Abstract: This paper introduces DW-Bench, a new benchmark that evaluates large language models (LLMs) on graph-topology reasoning over data warehouse schemas, exp

EgoMotion: Hierarchical Reasoning and Diffusion for Egocentric Vision-Language Motion Generation

ResearchDGX agent

arXiv:2604.19105v1 Announce Type: new Abstract: Faithfully modeling human behavior in dynamic environments is a foundational challenge for embodied intelligence. While conditional motion synthesis has

Enforcing Reciprocity in Operator Learning for Seismic Wave Propagation

ResearchDGX agent

arXiv:2602.11631v2 Announce Type: replace-cross Abstract: Accurate and efficient wavefield modeling underpins seismic structure and source studies. Traditional methods comply with physical laws but ar

Enjoyed the read? If you have deep experience in ML frameworks (training or inference) and love working on problems like these, our team is …

Model ReleasesDGX agent

Enjoyed the read? If you have deep experience in ML frameworks (training or inference) and love working on problems like these, our team is hiring! ML Systems Engineer, Frameworks & Tooling: https://j

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation

ApplicationsDGX agent

arXiv:2604.19678v1 Announce Type: new Abstract: Function vectors (FVs) are vector representations of tasks extracted from model activations during in-context learning. While prior work has shown that

Fast and Robust Diffusion Posterior Sampling for MR Image Reconstruction Using the Preconditioned Unadjusted Langevin Algorithm

Model ReleasesDGX agent

arXiv:2512.05791v2 Announce Type: replace-cross Abstract: Purpose: The Unadjusted Langevin Algorithm (ULA) in combination with diffusion models can generate high quality MRI reconstructions with uncer

Flux 2-Klein-9B NVFP4 works well on my RTX 3050, but it takes 55sec to generate 1024 resolution.

Local AiDGX agent

Flux 2-Klein-9B NVFP4 is a quantized image generation model that runs on mid-range GPUs like the RTX 3050. On this hardware, the model produces 1024-resolution images but with relatively slow inferenc

FoNE: Precise Single-Token Number Embeddings via Fourier Features

TutorialsDGX agent

arXiv:2502.09741v2 Announce Type: replace Abstract: Large Language Models (LLMs) typically represent numbers using multiple tokens, which requires the model to aggregate these tokens to interpret nume

From Natural Language to Executable Narsese: A Neuro-Symbolic Benchmark and Pipeline for Reasoning with NARS

Model ReleasesDGX agent

arXiv:2604.18873v1 Announce Type: new Abstract: Large language models (LLMs) are highly capable at language generation, but they remain unreliable when reasoning requires explicit symbolic structure,

Google unveils two new TPUs designed for the 'agentic era'

AgentsDGX agent

Google unveiled two eighth-generation Tensor Processing Units—the TPU 8t for training and TPU 8i for inference—designed with specialized architectures for AI model training and agent development. The

Has Automated Essay Scoring Reached Sufficient Accuracy? Deriving Achievable QWK Ceilings from Classical Test Theory

Model ReleasesDGX agent

arXiv:2604.19131v1 Announce Type: new Abstract: Automated essay scoring (AES) is commonly evaluated on public benchmarks using quadratic weighted kappa (QWK). However, because benchmark labels are ass

https://x.com/tanayj/status/2046733145446502805?s=46

ToolsDGX agent

https://x.com/tanayj/status/2046733145446502805?s=46 My read on the structure of this deal: 1. xAI is leveraging Cursor's data / traces to help train a better coding model (both Grok base model and po

KG-ViP: Bridging Knowledge Grounding and Visual Perception in Multi-modal LLMs for Visual Question Answering

ResearchDGX agent

arXiv:2601.11632v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) for Visual Question Answering (VQA) often suffer from dual limitations: knowledge hallucination and insuff

Knowledge-Guided Time-Varying Causal Inference for Arctic Sea Ice Dynamics

SafetyDGX agent

arXiv:2601.17647v2 Announce Type: replace-cross Abstract: Quantifying the causal relationship between sea ice thickness and sea surface height (SSH) is essential for understanding the mechanisms drivi

Latent Linear Quadratic Regulator for Robotic Control Tasks

TutorialsDGX agent

arXiv:2407.11107v3 Announce Type: replace-cross Abstract: Model predictive control (MPC) has played a more crucial role in various robotic control tasks, but its high computational requirements are co

Making ChatGPT better for clinicians

Model ReleasesDGX agent

OpenAI describes improvements and adaptations made to ChatGPT to better serve clinical and healthcare professionals. The work likely addresses how the model can be made more reliable, accurate, and us

Maximizing Gemini: Google Cloud makes its bid to build the operating system for enterprise AI

Model ReleasesDGX agent

Google LLC has emerged as the only cloud “hyperscaler” with a leading frontier artificial intelligence large language model – Gemini – and today it issued a raft of announcements designed to capitaliz

ODMA: On-Demand Memory Allocation Strategy for LLM Serving on LPDDR-Class Accelerators

Model ReleasesDGX agent

arXiv:2512.09427v5 Announce Type: replace-cross Abstract: Existing memory management techniques severely hinder efficient Large Language Model serving on accelerators constrained by poor random-access

Prompt to Pwn: Automated Exploit Generation for Smart Contracts

ApplicationsDGX agent

arXiv:2508.01371v3 Announce Type: replace-cross Abstract: Smart contracts are important for digital finance, yet they are hard to patch once deployed. Prior work has mainly explored LLMs for smart con

Qwen3.6 27B is now in LM Studio! Vision, reasoning, and agentic tool calling - running locally on your computer. Surpasses previous Qwen mod…

Model ReleasesDGX agent

Qwen3.6 27B is now in LM Studio! Vision, reasoning, and agentic tool calling - running locally on your computer. Surpasses previous Qwen models many times its size 🚀👾🔥 https://lmstudio.ai/models/qwen/

← Previous
1…425426427428429…1059
Next →