AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,157 results
22 Apr 2026

🔥 Emad Mostaque on Doom Debates 🔥 @EMostaque helped kick off the modern AI revolution as the head of Stability AI, the company behind Stab…

IndustryDGX agent

🔥 Emad Mostaque on Doom Debates 🔥 @EMostaque helped kick off the modern AI revolution as the head of Stability AI, the company behind Stable Diffusion. But unlike most AI CEOs, he doesn't sugarcoat th

From Natural Language to Executable Narsese: A Neuro-Symbolic Benchmark and Pipeline for Reasoning with NARS

Model ReleasesDGX agent

arXiv:2604.18873v1 Announce Type: new Abstract: Large language models (LLMs) are highly capable at language generation, but they remain unreliable when reasoning requires explicit symbolic structure,

I kinda think it would be interesting to have a structure where founders take 2 and 20 2% management fee on capital raised for running the c…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Industry
DGX agent

I kinda think it would be interesting to have a structure where founders take 2 and 20 2% management fee on capital raised for running the company 20% of any exit Single magic share with lots of votin

Let me say this clearly: LLMs cannot feel emotions. Emotions are evolutionary mechanisms. They push us to avoid danger or approach what is b…

SafetyDGX agent

Let me say this clearly: LLMs cannot feel emotions. Emotions are evolutionary mechanisms. They push us to avoid danger or approach what is beneficial. We experience emotions because we are alive, and

Multi-Cycle Spatio-Temporal Adaptation in Human-Robot Teaming

TutorialsDGX agent

arXiv:2604.19670v1 Announce Type: cross Abstract: Effective human-robot teaming is crucial for the practical deployment of robots in human workspaces. However, optimizing joint human-robot plans remai

OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration

AgentsDGX agent

arXiv:2505.11765v3 Announce Type: replace-cross Abstract: Agents powered by advanced large language models (LLMs) have demonstrated impressive capabilities across diverse complex applications. Recentl

Optimal Routing for Federated Learning over Dynamic Satellite Networks: Tractable or Not?

Local AiDGX agent

arXiv:2604.19399v1 Announce Type: new Abstract: Federated learning (FL) is a key paradigm for distributed model learning across decentralized data sources. Communication in each FL round typically con

Regulating Artificial Intimacy: From Locks and Blocks to Relational Accountability

SafetyDGX agent

arXiv:2604.18893v1 Announce Type: cross Abstract: A series of high-profile tragedies involving companion chatbots has triggered an unusually rapid regulatory response. Several jurisdictions, including

Revisiting RaBitQ and TurboQuant: A Symmetric Comparison of Methods, Theory, and Experiments

Model ReleasesDGX agent

arXiv:2604.19528v1 Announce Type: cross Abstract: This technical note revisits the relationship between RaBitQ and TurboQuant under a unified comparison framework. We compare the two methods in terms

RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian

ApplicationsDGX agent

arXiv:2604.19593v1 Announce Type: cross Abstract: The importance of clear and correct text in legal documents cannot be understated, and, consequently, a grammatical error correction tool meant to ass

Source: https://news.mit.edu/2026/new-type-electrically-driven-artificial-muscle-fiber-0409?utm_source=robotnews.therundown.ai&utm_medium=ne…

IndustryDGX agent

Source: https://news.mit.edu/2026/new-type-electrically-driven-artificial-muscle-fiber-0409?utm_source=robotnews.therundown.ai&utm_medium=newsletter&utm_campaign=unitree-s-cheapest-humanoid-goes-globa

The great AI democratization begins. Every startup just got a PhD-level ML team for free 🧵

IndustryDGX agent

This thread discusses how advances in AI tooling and accessible models have lowered barriers to entry for startups, enabling small teams to leverage capabilities previously requiring specialized ML ex

21 Apr 2026

CHIMERA: A Knowledge Base of Scientific Idea Recombinations for Research Analysis and Ideation

ResearchDGX agent

arXiv:2505.20779v5 Announce Type: replace Abstract: A hallmark of human innovation is recombination -- the creation of novel ideas by integrating elements from existing concepts and mechanisms. In thi

COFFAIL: A Dataset of Successful and Anomalous Robot Skill Executions in the Context of Coffee Preparation

SafetyDGX agent

arXiv:2604.18236v1 Announce Type: new Abstract: In the context of robot learning for manipulation, curated datasets are an important resource for advancing the state of the art; however, available dat

Common Corpus: The Largest Collection of Ethical Data for LLM Pre-Training

ApplicationsDGX agent

arXiv:2506.01732v2 Announce Type: replace Abstract: Large Language Models (LLMs) are pre-trained on large data from different sources and domains. These datasets often contain trillions of tokens, inc

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models

ApplicationsDGX agent

arXiv:2603.04592v3 Announce Type: replace Abstract: Standard Large Language Models (LLMs) are predominantly designed for static inference with pre-defined inputs, which limits their applicability in d

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

HardwareDGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

Matlas: A Semantic Search Engine for Mathematics

ResearchDGX agent

arXiv:2604.17484v1 Announce Type: cross Abstract: Retrieving mathematical knowledge is a central task in both human-driven research, such as determining whether a result already exists, finding relate

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework

AgentsDGX agent

arXiv:2511.21686v2 Announce Type: replace Abstract: Synthetic data has become increasingly important for training large language models, especially when real data is scarce, expensive, or privacy-sens

ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition

Model ReleasesDGX agent

arXiv:2503.21248v3 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in assisting scientific research, yet their ability to discover high-quality research hypotheses r

R&F-Inventory: A Large-Scale Dataset for Monotonic Inventory Estimation in Reach and Frequency Advertising

Model ReleasesDGX agent

arXiv:2604.16821v1 Announce Type: new Abstract: Reach and Frequency (R&F) contract advertising is an important form of widely used brand advertising. Unlike performance advertising, R&F contracts emph

SciImpact: A Multi-Dimensional, Multi-Field Benchmark for Scientific Impact Prediction

Model ReleasesDGX agent

arXiv:2604.17141v1 Announce Type: new Abstract: The rapid growth of scientific literature calls for automated methods to assess and predict research impact. Prior work has largely focused on citation-

SeekerGym: A Benchmark for Reliable Information Seeking

Model ReleasesDGX agent

arXiv:2604.17143v1 Announce Type: new Abstract: Despite their substantial successes, AI agents continue to face fundamental challenges in terms of trustworthiness. Consider deep research agents, taske

Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation

Local AiDGX agent

arXiv:2604.17656v1 Announce Type: cross Abstract: Video-to-music (V2M) is the fundamental task of creating background music for an input video. Recent V2M models achieve audiovisual alignment by typic

What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations

AgentsDGX agent

arXiv:2510.17795v3 Announce Type: replace Abstract: Replicating AI research is a crucial yet challenging task for large language model (LLM) agents. Existing approaches often struggle to generate exec

20 Apr 2026

Eco-Bee: A Personalised Multi-Modal Agent for Advancing Student Climate Awareness and Sustainable Behaviour in Campus Ecosystems

AgentsDGX agent

arXiv:2604.15327v1 Announce Type: cross Abstract: Universities are microcosms of urban ecosystems, with concentrated consumption patterns in food, transport, energy, and product usage. These environme

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

Model ReleasesDGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

SafetyDGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

We are happy to share early results from Logos, our novel first-principles augmented intelligence system, that has enabled insightful result…

IndustryDGX agent

We are happy to share early results from Logos, our novel first-principles augmented intelligence system, that has enabled insightful results across domains. We start with series of results in physics

19 Apr 2026

http://Localmaxxing.com is in private testing right now Looking to release public this week No longer do you have to post benchmarks into th…

IndustryDGX agent

Localmaxxing.com is a platform in private testing phase with plans for public release the same week this announcement was made, designed to simplify benchmark sharing by eliminating the need for users

I am not convinced that we should be comfortable calling 'problem solving' or 'judgement' or whatever as skills that are impossible for AI t…

ApplicationsDGX agent

I am not convinced that we should be comfortable calling 'problem solving' or 'judgement' or whatever as skills that are impossible for AI to do well. Like any other skill, there are humans who are re

We only have spotty information about this very important topic. It suggests AI can be good at diagnosis, but the real world doesn't always …

ApplicationsDGX agent

We only have spotty information about this very important topic. It suggests AI can be good at diagnosis, but the real world doesn't always match the experiments. https://x.com/emollick/status/1980474

WildChat is an amazing project. The project page seems not to be updated anymore, so people might not know that the dataset has been updated…

IndustryDGX agent

WildChat is an amazing project. The project page seems not to be updated anymore, so people might not know that the dataset has been updated with a very large number of transcripts through July 2025:

18 Apr 2026

One of the premier journals in my field... I think there are very valid reasons to set rules on AI in peer review (including disclosure), bu…

ApplicationsDGX agent

One of the premier journals in my field... I think there are very valid reasons to set rules on AI in peer review (including disclosure), but the idea that all AI models steal your data is very 2023.

This is far from an exhaustive list - look at my past tweets to find dozens more academics doing interesting work on the topic

ApplicationsDGX agent

This post references a non-exhaustive collection of academics working on a particular topic, with Ethan Mollick directing readers to his previous tweets for additional examples and research. The speci

17 Apr 2026

Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning

ApplicationsDGX agent

arXiv:2604.14161v1 Announce Type: new Abstract: Reliable evaluation is essential in machine learning research, yet methodological flaws-particularly data leakage-continue to undermine the validity of

DeepPresenter: Environment-Grounded Reflection for Agentic Presentation Generation

AgentsDGX agent

arXiv:2602.22839v2 Announce Type: replace Abstract: Presentation generation requires deep content research, coherent visual design, and iterative refinement based on observation. However, existing pre

Low-Cost System for Automatic Recognition of Driving Pattern in Assessing Interurban Mobility using Geo-Information

ResearchDGX agent

arXiv:2604.15216v1 Announce Type: cross Abstract: Mobility in urban and interurban areas, mainly by cars, is a day-to-day activity of many people. However, some of its main drawbacks are traffic jams

Quantitative Approximation Rates for Group Equivariant Learning

ResearchDGX agent

arXiv:2602.20370v2 Announce Type: replace Abstract: The universal approximation theorem establishes that neural networks can approximate any continuous function on a compact set. Later works in approx

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

Model ReleasesDGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

16 Apr 2026

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, …

Model ReleasesDGX agent

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, and retrospective curation. Production work is messier, with

Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems

SafetyDGX agent

arXiv:2604.13079v1 Announce Type: cross Abstract: Current AI alignment paradigms rely on behavioral correction: external supervisors (e.g., RLHF) observe outputs, judge against preferences, and adjust

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

Model ReleasesDGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

Model ReleasesDGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

Model ReleasesDGX agent

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

Hybrid Approach for Enhancing Lesion Segmentation in Fundus Images

ResearchDGX agent

arXiv:2509.25549v2 Announce Type: replace Abstract: Choroidal nevi are common benign pigmented lesions in the eye, with a small risk of transforming into melanoma. Early detection is critical to impro

I have found that asking for a sestina regularly triggers Opus 4.7's safety guardrails. The forbidden poetic form!

SafetyDGX agent

Ethan Mollick reported that requesting Claude Opus 4.7 to write sestinas—a complex poetic form with strict structural requirements—frequently triggers the model's safety guardrails, suggesting the AI

Its noticeable how much of the whole practice of working with AI - the prompts, the skill files, the connectors, retrieval work, the markdow…

ApplicationsDGX agent

Its noticeable how much of the whole practice of working with AI - the prompts, the skill files, the connectors, retrieval work, the markdown files, etc. - is a substitute for the real problem of cont

Just launched some great new providers in Stripe Projects (http://projects.dev): @huggingface @Cloudflare @OpenRouter @firecrawl @flydotio @…

IndustryDGX agent

Just launched some great new providers in Stripe Projects (http://projects.dev): @huggingface @Cloudflare @OpenRouter @firecrawl @flydotio @Amplitude_HQ @mixpanel @inngest Many more coming later this

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

SafetyDGX agent

arXiv:2601.15170v2 Announce Type: replace Abstract: The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult t

NEW Research from Google. Integration test failures are painful because the signal is buried in messy logs. Massive output, heterogeneous sy…

TutorialsDGX agent

NEW Research from Google. Integration test failures are painful because the signal is buried in messy logs. Massive output, heterogeneous systems, low signal-to-noise ratio, and unclear root causes. T

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

Model ReleasesDGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

The open-source AI community just got a new home for their data workflows. 🤗 @huggingface is now available in Adaptive Data. Pull datasets …

IndustryDGX agent

The open-source AI community just got a new home for their data workflows. 🤗 @huggingface is now available in Adaptive Data. Pull datasets directly into a platform that evolves with the problems you'r

Why having “humans in the loop” in an AI war is an illusion

ApplicationsDGX agent

The availability of artificial intelligence for use in warfare is at the center of a legal battle between Anthropic and the Pentagon. This debate has become urgent, with AI playing a bigger role than

15 Apr 2026

Beyond Single-Dimension Novelty: How Combinations of Theory, Method, and Results-based Novelty Shape Scientific Impact

Model ReleasesDGX agent

arXiv:2604.12471v1 Announce Type: cross Abstract: Scientific novelty drives advances at the research frontier, yet it is also associated with heightened uncertainty and potential resistance from incum

Calibrated Confidence Estimation for Tabular Question Answering

ResearchDGX agent

arXiv:2604.12491v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for tabular question answering, yet calibration on structured data is largely unstudied. This pap

Deep Learning using Rectified Linear Units (ReLU)

ResearchDGX agent

arXiv:1803.08375v3 Announce Type: replace-cross Abstract: The Rectified Linear Unit (ReLU) is a foundational activation function in artficial neural networks. Recent literature frequently misattribute

Document OCR benchmarks are still an open problem Existing document OCR benchmarks are either too narrowly focused on a specific type (e.g. …

Model ReleasesDGX agent

Document OCR benchmarks are still an open problem Existing document OCR benchmarks are either too narrowly focused on a specific type (e.g. FinTabNet, ChartQA), or on documents that aren’t reflective

DRPG (Decompose, Retrieve, Plan, Generate): An Agentic Framework for Academic Rebuttal

AgentsDGX agent

arXiv:2601.18081v2 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in scientific research workflows, automated support for academic rebuttal, a crucial st

Efficient and Scalable Granular-ball Graph Coarsening Method for Large-scale Graph Node Classification

ResearchDGX agent

arXiv:2603.29148v2 Announce Type: replace-cross Abstract: Graph Convolutional Network (GCN) is a model that can effectively handle graph data tasks and has been successfully applied. However, for larg

← Previous
1…4950515253…203
Next →