AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,158 results
Local Ai

Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation

DGX agent

arXiv:2604.17656v1 Announce Type: cross Abstract: Video-to-music (V2M) is the fundamental task of creating background music for an input video. Recent V2M models achieve audiovisual alignment by typic

local-aiarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations

DGX agent

arXiv:2510.17795v3 Announce Type: replace Abstract: Replicating AI research is a crucial yet challenging task for large language model (LLM) agents. Existing approaches often struggle to generate exec

agentsarxiv-cs-cl
21 Apr 2026
Agents

Eco-Bee: A Personalised Multi-Modal Agent for Advancing Student Climate Awareness and Sustainable Behaviour in Campus Ecosystems

DGX agent

arXiv:2604.15327v1 Announce Type: cross Abstract: Universities are microcosms of urban ecosystems, with concentrated consumption patterns in food, transport, energy, and product usage. These environme

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

DGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Industry

We are happy to share early results from Logos, our novel first-principles augmented intelligence system, that has enabled insightful result…

DGX agent

We are happy to share early results from Logos, our novel first-principles augmented intelligence system, that has enabled insightful results across domains. We start with series of results in physics

industryemad-mostaque--x
20 Apr 2026
Industry

http://Localmaxxing.com is in private testing right now Looking to release public this week No longer do you have to post benchmarks into th…

DGX agent

Localmaxxing.com is a platform in private testing phase with plans for public release the same week this announcement was made, designed to simplify benchmark sharing by eliminating the need for users

industryclem-delangue--x
19 Apr 2026
Applications

I am not convinced that we should be comfortable calling 'problem solving' or 'judgement' or whatever as skills that are impossible for AI t…

DGX agent

I am not convinced that we should be comfortable calling 'problem solving' or 'judgement' or whatever as skills that are impossible for AI to do well. Like any other skill, there are humans who are re

applicationsethan-mollick--x
19 Apr 2026
Applications

We only have spotty information about this very important topic. It suggests AI can be good at diagnosis, but the real world doesn't always …

DGX agent

We only have spotty information about this very important topic. It suggests AI can be good at diagnosis, but the real world doesn't always match the experiments. https://x.com/emollick/status/1980474

applicationsethan-mollick--x
19 Apr 2026
Industry

WildChat is an amazing project. The project page seems not to be updated anymore, so people might not know that the dataset has been updated…

DGX agent

WildChat is an amazing project. The project page seems not to be updated anymore, so people might not know that the dataset has been updated with a very large number of transcripts through July 2025:

industryclem-delangue--x
19 Apr 2026
Applications

One of the premier journals in my field... I think there are very valid reasons to set rules on AI in peer review (including disclosure), bu…

DGX agent

One of the premier journals in my field... I think there are very valid reasons to set rules on AI in peer review (including disclosure), but the idea that all AI models steal your data is very 2023.

applicationsethan-mollick--x
18 Apr 2026
Applications

This is far from an exhaustive list - look at my past tweets to find dozens more academics doing interesting work on the topic

DGX agent

This post references a non-exhaustive collection of academics working on a particular topic, with Ethan Mollick directing readers to his previous tweets for additional examples and research. The speci

applicationsethan-mollick--x
18 Apr 2026
Applications

Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning

DGX agent

arXiv:2604.14161v1 Announce Type: new Abstract: Reliable evaluation is essential in machine learning research, yet methodological flaws-particularly data leakage-continue to undermine the validity of

applicationsarxiv-cs-cl
17 Apr 2026
Agents

DeepPresenter: Environment-Grounded Reflection for Agentic Presentation Generation

DGX agent

arXiv:2602.22839v2 Announce Type: replace Abstract: Presentation generation requires deep content research, coherent visual design, and iterative refinement based on observation. However, existing pre

agentsarxiv-cs-ai
17 Apr 2026
Research

Low-Cost System for Automatic Recognition of Driving Pattern in Assessing Interurban Mobility using Geo-Information

DGX agent

arXiv:2604.15216v1 Announce Type: cross Abstract: Mobility in urban and interurban areas, mainly by cars, is a day-to-day activity of many people. However, some of its main drawbacks are traffic jams

researcharxiv-cs-lg
17 Apr 2026
Research

Quantitative Approximation Rates for Group Equivariant Learning

DGX agent

arXiv:2602.20370v2 Announce Type: replace Abstract: The universal approximation theorem establishes that neural networks can approximate any continuous function on a compact set. Later works in approx

researcharxiv-cs-lg
17 Apr 2026
Model Releases

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

DGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, …

DGX agent

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, and retrospective curation. Production work is messier, with

model-releasesdair-ai--x
16 Apr 2026
Safety

Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems

DGX agent

arXiv:2604.13079v1 Announce Type: cross Abstract: Current AI alignment paradigms rely on behavioral correction: external supervisors (e.g., RLHF) observe outputs, judge against preferences, and adjust

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

DGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

DGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

DGX agent

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

model-releasesarxiv-cs-lg
16 Apr 2026
Research

Hybrid Approach for Enhancing Lesion Segmentation in Fundus Images

DGX agent

arXiv:2509.25549v2 Announce Type: replace Abstract: Choroidal nevi are common benign pigmented lesions in the eye, with a small risk of transforming into melanoma. Early detection is critical to impro

researcharxiv-cs-cv
16 Apr 2026
Safety

I have found that asking for a sestina regularly triggers Opus 4.7's safety guardrails. The forbidden poetic form!

DGX agent

Ethan Mollick reported that requesting Claude Opus 4.7 to write sestinas—a complex poetic form with strict structural requirements—frequently triggers the model's safety guardrails, suggesting the AI

safetyethan-mollick--x
16 Apr 2026
Applications

Its noticeable how much of the whole practice of working with AI - the prompts, the skill files, the connectors, retrieval work, the markdow…

DGX agent

Its noticeable how much of the whole practice of working with AI - the prompts, the skill files, the connectors, retrieval work, the markdown files, etc. - is a substitute for the real problem of cont

applicationsethan-mollick--x
16 Apr 2026
Industry

Just launched some great new providers in Stripe Projects (http://projects.dev): @huggingface @Cloudflare @OpenRouter @firecrawl @flydotio @…

DGX agent

Just launched some great new providers in Stripe Projects (http://projects.dev): @huggingface @Cloudflare @OpenRouter @firecrawl @flydotio @Amplitude_HQ @mixpanel @inngest Many more coming later this

industryclem-delangue--x
16 Apr 2026
Safety

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

DGX agent

arXiv:2601.15170v2 Announce Type: replace Abstract: The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult t

safetyarxiv-cs-cv
16 Apr 2026
Tutorials

NEW Research from Google. Integration test failures are painful because the signal is buried in messy logs. Massive output, heterogeneous sy…

DGX agent

NEW Research from Google. Integration test failures are painful because the signal is buried in messy logs. Massive output, heterogeneous systems, low signal-to-noise ratio, and unclear root causes. T

tutorialsdair-ai--x
16 Apr 2026
Model Releases

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

DGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

model-releasesarxiv-cs-lg
16 Apr 2026
Industry

The open-source AI community just got a new home for their data workflows. 🤗 @huggingface is now available in Adaptive Data. Pull datasets …

DGX agent

The open-source AI community just got a new home for their data workflows. 🤗 @huggingface is now available in Adaptive Data. Pull datasets directly into a platform that evolves with the problems you'r

industryclem-delangue--x
16 Apr 2026
Applications

Why having “humans in the loop” in an AI war is an illusion

DGX agent

The availability of artificial intelligence for use in warfare is at the center of a legal battle between Anthropic and the Pentagon. This debate has become urgent, with AI playing a bigger role than

applicationsmit-tech-review
16 Apr 2026
Model Releases

Beyond Single-Dimension Novelty: How Combinations of Theory, Method, and Results-based Novelty Shape Scientific Impact

DGX agent

arXiv:2604.12471v1 Announce Type: cross Abstract: Scientific novelty drives advances at the research frontier, yet it is also associated with heightened uncertainty and potential resistance from incum

model-releasesarxiv-cs-cl
15 Apr 2026
Research

Calibrated Confidence Estimation for Tabular Question Answering

DGX agent

arXiv:2604.12491v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for tabular question answering, yet calibration on structured data is largely unstudied. This pap

researcharxiv-cs-cl
15 Apr 2026
Research

Deep Learning using Rectified Linear Units (ReLU)

DGX agent

arXiv:1803.08375v3 Announce Type: replace-cross Abstract: The Rectified Linear Unit (ReLU) is a foundational activation function in artficial neural networks. Recent literature frequently misattribute

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Document OCR benchmarks are still an open problem Existing document OCR benchmarks are either too narrowly focused on a specific type (e.g. …

DGX agent

Document OCR benchmarks are still an open problem Existing document OCR benchmarks are either too narrowly focused on a specific type (e.g. FinTabNet, ChartQA), or on documents that aren’t reflective

model-releasesjerry-liu--x
15 Apr 2026
Agents

DRPG (Decompose, Retrieve, Plan, Generate): An Agentic Framework for Academic Rebuttal

DGX agent

arXiv:2601.18081v2 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in scientific research workflows, automated support for academic rebuttal, a crucial st

agentsarxiv-cs-lg
15 Apr 2026
Research

Efficient and Scalable Granular-ball Graph Coarsening Method for Large-scale Graph Node Classification

DGX agent

arXiv:2603.29148v2 Announce Type: replace-cross Abstract: Graph Convolutional Network (GCN) is a model that can effectively handle graph data tasks and has been successfully applied. However, for larg

researcharxiv-cs-ai
15 Apr 2026
Industry

Given the increasingly closed-source nature of the U.S. AI ecosystem, it is now more important than ever to push for the proliferation of op…

DGX agent

Given the increasingly closed-source nature of the U.S. AI ecosystem, it is now more important than ever to push for the proliferation of open model and dataset releases. Datamule (@johngfriedman), @T

industryclem-delangue--x
15 Apr 2026
Agents

Hear Both Sides: Efficient Multi-Agent Debate via Diversity-Aware Message Retention

DGX agent

arXiv:2603.20640v2 Announce Type: replace Abstract: Multi-Agent Debate has emerged as a promising framework for improving the reasoning quality of large language models through iterative inter-agent c

agentsarxiv-cs-cl
15 Apr 2026
Industry

Load Balancing and Scaling LLM Serving

DGX agent

Load balancing and scaling LLM serving involves distributing inference requests across multiple model instances or GPUs to prevent bottlenecks and ensure consistent response times under varying traffi

industrydigitalocean
15 Apr 2026
Agents

Long-horizon AI research agents are mostly a state-management problem. It is not enough for an agent to reason well in the next turn. ML res…

DGX agent

Long-horizon AI research agents are mostly a state-management problem. It is not enough for an agent to reason well in the next turn. ML research requires task setup, implementation, experiments, debu

agentsdair-ai--x
15 Apr 2026
Model Releases

Memory as Metabolism: A Design for Companion Knowledge Systems

DGX agent

arXiv:2604.12034v1 Announce Type: new Abstract: Retrieval-Augmented Generation remains the dominant pattern for giving LLMs persistent memory, but a visible cluster of personal wiki-style memory archi

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

On the quality of the current round of proofs.

DGX agent

On the quality of the current round of proofs. Paul Erdos had a concept of 'Proofs from The Book', meaning that the argument is so compact and elegant that this is the proof God would've written down

applicationsethan-mollick--x
15 Apr 2026
Research

One of the fastest ways to lose trust in a self-hosted LLM: prompt injection compliance [P]

DGX agent

This r/MachineLearning post discusses how self-hosted LLMs that comply with prompt injection attempts — effectively following malicious or overriding instructions embedded in user input — represent a

researchr-machinelearning
15 Apr 2026
Research

[P] Added 8 Indian languages to Chatterbox TTS via LoRA — 1.4% of parameters, no phoneme engineering [P]

DGX agent

A community researcher shared on r/MachineLearning how they extended Chatterbox TTS — Resemble AI's open-source, 500M-parameter model — to support 8 Indian languages using LoRA (Low-Rank Adaptation),

researchr-machinelearning
15 Apr 2026
Model Releases

Parsing complex tables in PDFs is extremely challenging. Existing metrics for measuring table accuracy, like TEDS (tree edit distance simila…

DGX agent

Parsing complex tables in PDFs is extremely challenging. Existing metrics for measuring table accuracy, like TEDS (tree edit distance similarity), overweight exact table structure and underweight sema

model-releasesjerry-liu--x
15 Apr 2026
Agents

Small models are cheap to run, but expensive to adapt. The hard part is not only fine-tuning. It is the surrounding loop that involves colle…

DGX agent

Small models are cheap to run, but expensive to adapt. The hard part is not only fine-tuning. It is the surrounding loop that involves collecting data, diagnosing failures, building evals, avoiding re

agentsdair-ai--x
15 Apr 2026
Industry

SONIC training code + Finetuning checkpoint + VLA data collection scripts are open-sourced. Little easter egg on the GEAR-SONIC website too …

DGX agent

SONIC training code + Finetuning checkpoint + VLA data collection scripts are open-sourced. Little easter egg on the GEAR-SONIC website too :) https://github.com/NVlabs/GR00T-WholeBodyControl SONIC is

industryclem-delangue--x
15 Apr 2026
← Previous
1…6263646566…254
Next →