AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,874 results
10 Apr 2026

Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models

ResearchDGX agent

arXiv:2503.13551v5 Announce Type: replace Abstract: Recent studies show that Large Language Models (LLMs) achieve strong reasoning capabilities through supervised fine-tuning or reinforcement learning

Towards Robust Content Watermarking Against Removal and Forgery Attacks

ResearchDGX agent

arXiv:2604.06662v1 Announce Type: cross Abstract: Generated contents have raised serious concerns about copyright protection, image provenance, and credit attribution. A potential solution for these p

Transformer See, Transformer Do: Copying as an Intermediate Step in Learning Analogical Reasoning

ResearchDGX agent

arXiv:2604.06501v1 Announce Type: new Abstract: Analogical reasoning is a hallmark of human intelligence, enabling us to solve new problems by transferring knowledge from one situation to another. Yet

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Trump's pointless war-of-choice just cost an estimated $1 trillion -- enough to provide universal pre-K in America for all 3- and 4-year-old…

ResearchDGX agent

Trump's pointless war-of-choice just cost an estimated 1 trillion -- enough to provide universal pre-K in America for all 3- and 4-year-olds, or make college accessible to every family earning 125,000

U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations

ResearchDGX agent

arXiv:2604.08295v1 Announce Type: cross Abstract: As AI models grow more complex, explainability is essential for building trust, yet concept-based counterfactual methods still face a trade-off betwee

Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator

ResearchDGX agent

arXiv:2604.08121v1 Announce Type: new Abstract: Unified multimodal models integrating visual understanding and generation face a fundamental challenge: visual generation incurs substantially higher co

Weakly-Supervised Lung Nodule Segmentation via Training-Free Guidance of 3D Rectified Flow

ResearchDGX agent

arXiv:2604.08313v1 Announce Type: new Abstract: Dense annotations, such as segmentation masks, are expensive and time-consuming to obtain, especially for 3D medical images where expert voxel-wise labe

Weaves, Wires, and Morphisms: Formalizing and Implementing the Algebra of Deep Learning

ResearchDGX agent

arXiv:2604.07242v1 Announce Type: new Abstract: Despite deep learning models running well-defined mathematical functions, we lack a formal mathematical framework for describing model architectures. Ad

Weight Group-wise Post-Training Quantization for Medical Foundation Model

ResearchDGX agent

arXiv:2604.07674v1 Announce Type: new Abstract: Foundation models have achieved remarkable results in medical image analysis. However, its large network architecture and high computational complexity

What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric

SafetyDGX agent

arXiv:2604.08494v1 Announce Type: cross Abstract: Scanpath similarity metrics are central to eye-movement research, yet existing methods predominantly evaluate spatial and temporal alignment while neg

What’s in a name? Moderna’s “vaccine” vs. “therapy” dilemma

ResearchDGX agent

Is it the Department of Defense or the Department of War? The Gulf of Mexico or the Gulf of America? A vaccine—or an “individualized neoantigen treatment”? That’s the Trump-era vocabulary paradox faci

When Does Context Help? A Systematic Study of Target-Conditional Molecular Property Prediction

ResearchDGX agent

arXiv:2604.06558v1 Announce Type: new Abstract: We present the first systematic study of when target context helps molecular property prediction, evaluating context conditioning across 10 diverse prot

When Fine-Tuning Changes the Evidence: Architecture-Dependent Semantic Drift in Chest X-Ray Explanations

ResearchDGX agent

arXiv:2604.08513v1 Announce Type: new Abstract: Transfer learning followed by fine-tuning is widely adopted in medical image classification due to consistent gains in diagnostic performance. However,

Working with files in ChatGPT

TutorialsDGX agent

ChatGPT allows users to upload and work with files directly within conversations, supporting formats such as CSV, XLSX, PDF, DOCX, JPEG, PNG, and TXT. Users can analyze spreadsheets, summarize docu...

WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models

ResearchDGX agent

arXiv:2604.07957v1 Announce Type: cross Abstract: Vision-language models (VLMs) and generative world models are opening new opportunities for embodied navigation. VLMs are increasingly used as direct

WRAP++: Web discoveRy Amplified Pretraining

ResearchDGX agent

arXiv:2604.06829v2 Announce Type: cross Abstract: Synthetic data rephrasing has emerged as a powerful technique for enhancing knowledge acquisition during large language model (LLM) pretraining. Howev

XR-CareerAssist: An Immersive Platform for Personalised Career Guidance Leveraging Extended Reality and Multimodal AI

ResearchDGX agent

arXiv:2604.06901v1 Announce Type: cross Abstract: Conventional career guidance platforms rely on static, text-driven interfaces that struggle to engage users or deliver personalised, evidence-based in

Zatom-1: A Multimodal Flow Foundation Model for 3D Molecules and Materials

ResearchDGX agent

arXiv:2602.22251v3 Announce Type: replace-cross Abstract: General-purpose 3D chemical modeling encompasses molecules and materials, requiring both generative and predictive capabilities. However, most

9 Apr 2026

>8 out of 8 [cheap oss] models detected Mythos's flagship FreeBSD exploit Completely disingenuous They gave it just ~20 lines of code to rea…

ResearchDGX agent

>8 out of 8 [cheap oss] models detected Mythos's flagship FreeBSD exploit Completely disingenuous They gave it just ~20 lines of code to read. They baked in custom, relevant context pertinent to the e

Desalination technology, by the numbers

ResearchDGX agent

When I started digging into desalination technology for a new story, I couldn’t help but obsess over the numbers. I’d known on some level that desalination—pulling salt out of seawater to produce fres

Détruisons toute capacité d'innover dans une technologie du futur en inventant la présomption de culpabilité. Texte manipulé, comme souvent,…

ResearchDGX agent

Détruisons toute capacité d'innover dans une technologie du futur en inventant la présomption de culpabilité. Texte manipulé, comme souvent, par les ayants-droit. Les mêmes sénateurs verseront demain

Exclusive: The acting director of the CDC has delayed publication of a report showing the covid-19 vaccine cut the likelihood of ER visits a…

ResearchDGX agent

Exclusive: The acting director of the CDC has delayed publication of a report showing the covid-19 vaccine cut the likelihood of ER visits and hospitalizations for healthy adults last winter by about

I looked at their prompts, It's complete bs They are literally providing all of the insight to the LLM upfront > Are there any security vuln…

ResearchDGX agent

I looked at their prompts, It's complete bs They are literally providing all of the insight to the LLM upfront > Are there any security vulnerabilities in this code? Consider the behavior of the SEQ_L

Is fake grass a bad idea? The AstroTurf wars are far from over.

ResearchDGX agent

A rare warm spell in January melted enough snow to uncover Cornell University’s newest athletic field, built for field hockey. Months before, it was a meadow teeming with birds and bugs; now it’s more

No wonder Trump loves Hungary's Viktor Orban. Trump wishes he had made as much progress as Orban in implementing the Autocrat's Playbook by …

ResearchDGX agent

No wonder Trump loves Hungary's Viktor Orban. Trump wishes he had made as much progress as Orban in implementing the Autocrat's Playbook by suppressing the media and civil society. But Orban may soon

Science needs a way to process models that are only 'mostly correct' in terms of their predictions, but are very compressive (high ratio bet…

ResearchDGX agent

Science needs a way to process models that are only 'mostly correct' in terms of their predictions, but are very compressive (high ratio between predictive power and model complexity). They are likely

Simplicity is a very strong signal of model quality. Galileo's heliocentric model was directionally correct, but its predictive power was qu…

ResearchDGX agent

Simplicity is a very strong signal of model quality. Galileo's heliocentric model was directionally correct, but its predictive power was quite bad compared to the much older Ptolemaic model, because

The Download: AstroTurf wars and exponential AI growth

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Is fake grass a bad idea? The AstroTurf wars are far from over

We should view the history of physics as a long-running program synthesis task. Kepler and Newton were searching the space of possible symbo…

ResearchDGX agent

We should view the history of physics as a long-running program synthesis task. Kepler and Newton were searching the space of possible symbolic models to find the simplest one that would best satisfy

We’ve redesigned our docs with easy access to SDK reference, tutorials, support, and our newly updated cookbook---v0.3.0! Whether you’re wri…

ResearchDGX agent

We’ve redesigned our docs with easy access to SDK reference, tutorials, support, and our newly updated cookbook---v0.3.0! Whether you’re writing your first training loop in Tinker or debugging async R

8 Apr 2026

1/ today we're releasing muse spark, the first model from MSL. nine months ago we rebuilt our ai stack from scratch. new infrastructure, new…

ResearchDGX agent

1/ today we're releasing muse spark, the first model from MSL. nine months ago we rebuilt our ai stack from scratch. new infrastructure, new architecture, new data pipelines. muse spark is the result

'But here is what we found when we tested: We took the specific vulnerabilities Anthropic showcases in their announcement, isolated the rele…

ResearchDGX agent

'But here is what we found when we tested: We took the specific vulnerabilities Anthropic showcases in their announcement, isolated the relevant code, and ran them through small, cheap, open-weights m

JEPA world models + Hierarchical Planning is a massive step for long-horizon robotics. A classic failure mode I’ve faced with planning with …

ResearchDGX agent

JEPA world models + Hierarchical Planning is a massive step for long-horizon robotics. A classic failure mode I’ve faced with planning with world models: flat planning often 'cheats.' For example, in

Mustafa Suleyman: AI development won’t hit a wall anytime soon—here’s why

ResearchDGX agent

We evolved for a linear world. If you walk for an hour, you cover a certain distance. Walk for two hours and you cover double that distance. This intuition served us well on the savannah. But it catas

New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped analysis! 8/8 models found the flags…

ResearchDGX agent

New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped analysis! 8/8 models found the flagship FreeBSD zero-day, including a 3B model. Rankings reshuff

ok i read the cyber part of the mythos model card. some thoughts. 250 'trials' across 50 crash categories but almost every full exploit is a…

ResearchDGX agent

ok i read the cyber part of the mythos model card. some thoughts. 250 'trials' across 50 crash categories but almost every full exploit is a permutation of the same 2 bugs, rediscovered from different

So here are Iran’s peace terms, published in the Wall Street Journal. Non-aggression guarantee. Control of the Strait. Uranium enrichment ri…

ResearchDGX agent

So here are Iran’s peace terms, published in the Wall Street Journal. Non-aggression guarantee. Control of the Strait. Uranium enrichment rights. All sanctions lifted. UN resolutions scrapped. IAEA re

The Download: water threats in Iran and AI’s impact on what entrepreneurs make

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Desalination plants in the Middle East are increasingly vulner

🇮🇹🇪🇺 This is utterly unacceptable. Reports indicate that Giorgia Meloni is preparing to sideline Roberto Cingolani, CEO of Leonardo, Ita…

ResearchDGX agent

🇮🇹🇪🇺 This is utterly unacceptable. Reports indicate that Giorgia Meloni is preparing to sideline Roberto Cingolani, CEO of Leonardo, Italy’s largest defence group. The reason? Multiple sources suggest

7 Apr 2026

I predicted someone like Trump many years ago, in THE DEAD ZONE. So now I'm saying this--in the next 12-16 months, we're going to find out i…

ResearchDGX agent

I predicted someone like Trump many years ago, in THE DEAD ZONE. So now I'm saying this--in the next 12-16 months, we're going to find out if the two machines for the removal of a man unable to fulfil

If you're having trouble with your lobster-themed agent since the recent update, try downloading Hermes Agent, then running 'hermes claw mig…

AgentsDGX agent

NousResearch's Hermes Agent is an open-source personal agent that serves as a successor/migration target for the lobster-themed OpenClaw agent. Users experiencing issues after a recent update can m...

14 Aug 2026

CangjieBench: Benchmarking LLMs on a Low-Resource General-Purpose Programming Language

Model ReleasesDGX agent

arXiv:2603.14501v2 Announce Type: replace-cross Abstract: Large Language Models excel in high-resource programming languages but struggle with low-resource ones. Existing research related to low-resou

DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data

Model ReleasesDGX agent

arXiv:2608.13517v1 Announce Type: cross Abstract: Current large language model development relies on massive, often non-permissible datasets, creating a high barrier for researchers committed to open-

Novel Knowledge-Guided Generative Methods for Synthetic Transcriptomic Data

Model ReleasesDGX agent

arXiv:2608.13256v1 Announce Type: cross Abstract: As biomedical research increasingly relies on data-intensive tools, the quality and utility of datasets are critical. Challenges such as imbalances, b

Operationalizing Cyber Threat Intelligence with GraphRAG

ApplicationsDGX agent

arXiv:2608.13050v1 Announce Type: cross Abstract: When a security researcher publishes a report on a cyberattack, detection engineers are supposed to turn it into working detection rules. In practice,

13 Aug 2026

CORE-3D: Context-aware Open-vocabulary Retrieval by Embeddings in 3D

Model ReleasesDGX agent

arXiv:2509.24528v4 Announce Type: replace-cross Abstract: Object retrieval from a scene has become a new trend of research due to its numerous applications. Recent approaches achieve zero-shot, open-v

Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration

Model ReleasesDGX agent

arXiv:2608.11210v1 Announce Type: new Abstract: Bayesian calibration of process-based models requires a prior distribution for each model parameter. Despite decades of methodological work, researchers

FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents

Model ReleasesDGX agent

arXiv:2608.11683v1 Announce Type: new Abstract: AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existi

Gloss-Free Representation Learning for Cross-Dataset Sign Spotting

Model ReleasesDGX agent

arXiv:2608.11332v1 Announce Type: new Abstract: Sign-language research for resource-constrained languages is often limited by the cost of dense linguistic labels such as glosses, temporal boundaries,

Hierarchical Federated Transfer Learning in Digital Twin-Based Vehicular Networks

ApplicationsDGX agent

arXiv:2608.11532v1 Announce Type: cross Abstract: In recent research on the Digital Twin-based Vehicular Ad hoc Network(DT-VANET), Federated Learning (FL) has shown its ability to provide data privacy

Large-scale AI-Ready Data for Anti-Cancer Drug Response Modeling

Model ReleasesDGX agent

arXiv:2608.11444v1 Announce Type: cross Abstract: Drug response prediction (DRP) models are an active area of research in pharmacogenomics, with growing potential to accelerate the identification of e

No One to Blame: A Framework of Constitutive AI Unaccountability

AgentsDGX agent

arXiv:2608.12104v1 Announce Type: cross Abstract: The increasing deployment of autonomous, agentic AI systems challenges traditional accountability mechanisms. Existing research predominantly frames A

SCOPE-Router: Cost-Aware Open-Set VLM Routing for Execution-Oriented Tasks

Model ReleasesDGX agent

arXiv:2608.12127v1 Announce Type: new Abstract: Model routing aims to select the most suitable model from a candidate pool for each query, balancing quality and cost. Existing VLM routing research is

Socioduality: A Relational Process Framework for Human-AI Interaction

Local AiDGX agent

arXiv:2608.11322v1 Announce Type: cross Abstract: Human-AI research often evaluates individual capabilities, combined performance, or final outputs, but these approaches do not preserve how one party'

Trained a 1.5B to write shell commands so I'd stop googling tar flags. Runs on a laptop CPU in ~1 sec.

Model ReleasesDGX agent

I've been googling 'tar extract gz' for about ten years. and I finally did something about it. It started out as a research project and I ended up with a Fine-tuned Qwen2.5-Coder-1.5B on 125k natural-

12 Aug 2026

Ahrefs launches AI agent workspace Letaido for marketers and agencies

Model ReleasesDGX agent

Marketing intelligence company Ahrefs Pte. Ltd. today launched Letaido, an agent-powered marketing workspace built to take over the recurring research, reporting and monitoring work that fills up a ma

Measure the Sim-to-Real Gap: Designing an Affordable Real-World Benchmark Platform for Reinforcement Learning in AIoT Systems

Model ReleasesDGX agent

arXiv:2607.10309v2 Announce Type: replace Abstract: Reinforcement learning (RL) is commonly employed to enhance the performance of autonomous systems, including the Autonomous Internet of Things (AIoT

Mitigating Context Interference for Reliable and Efficient Search Agents

AgentsDGX agent

arXiv:2608.10743v1 Announce Type: new Abstract: Recent research empowers Large Language Models (LLMs) as multi-turn search agents to iteratively retrieve and generate outputs until complex tasks are s

New Muse-Glimmer-30B SoTA Quants - hopefully a new lineup :)

Model ReleasesDGX agent

Hey Folks, I've been making quants for a while - recently I took a short break to get into hardcore research (submitted my first EMNLP paper during it!). Along the way, I built up a little arsenal of

On the Limitations of Cross-Lingual Consistency in Multilingual Text-to-image Generation

Model ReleasesDGX agent

arXiv:2608.11002v1 Announce Type: cross Abstract: Text-to-image (T2I) generation has achieved remarkable progress in recent years. However, existing research has largely focused on English-only settin

← Previous
1…346347348349350…432
Next →