AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
17 Apr 2026

Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error

SafetyDGX agent

arXiv:2510.26109v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly boosted the reasoning capability of language models (LMs). However, existing

Emergent Neural Automaton Policies: Learning Symbolic Structure from Visuomotor Trajectories

SafetyDGX agent

arXiv:2603.25903v2 Announce Type: replace Abstract: Scaling robot learning to long-horizon tasks remains a formidable challenge. While end-to-end policies often lack the structural priors needed for e

Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.24120v2 Announce Type: replace Abstract: Automated neural network architecture design remains a significant challenge in computer vision. Task diversity and computational constraints requir

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

SafetyDGX agent

arXiv:2604.14267v1 Announce Type: new Abstract: Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavaila

Expressivity of Transformers: A Tropical Geometry Perspective

ResearchDGX agent

arXiv:2604.14727v1 Announce Type: new Abstract: To quantify the geometric expressivity of transformers, we introduce a tropical geometry framework to characterize their exact spatial partitioning capa

For those who made it this far, use code 'HERMESAGENT0010' for $10 off a valid Nous Portal Subscription (new or existing). Code valid for th…

AgentsDGX agent

For those who made it this far, use code 'HERMESAGENT0010' for $10 off a valid Nous Portal Subscription (new or existing). Code valid for the first 250 users. Subscribe to Nous Portal: http://portal.n

From Boundaries to Semantics: Prompt-Guided Multi-Task Learning for Petrographic Thin-section Segmentation

SafetyDGX agent

arXiv:2604.14805v1 Announce Type: new Abstract: Grain-edge segmentation (GES) and lithology semantic segmentation (LSS) are two pivotal tasks for quantifying rock fabric and composition. However, thes

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution

TutorialsDGX agent

arXiv:2604.15097v1 Announce Type: cross Abstract: This beta technical report asks how reusable experience should be represented so that it can function as effective test-time control and as a substrat

Fundamental Limitations of Favorable Privacy-Utility Guarantees for DP-SGD

ResearchDGX agent

arXiv:2601.10237v2 Announce Type: replace Abstract: Differentially Private Stochastic Gradient Descent (DP-SGD) is the dominant paradigm for private training, but its fundamental limitations under wor

Gating Enables Curvature: A Geometric Expressivity Gap in Attention

ResearchDGX agent

arXiv:2604.14702v1 Announce Type: new Abstract: Multiplicative gating is widely used in neural architectures and has recently been applied to attention layers to improve performance and training stabi

GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

SafetyDGX agent

arXiv:2604.14258v1 Announce Type: cross Abstract: Large language models are typically post-trained using supervised fine-tuning (SFT) and reinforcement learning (RL), yet effectively unifying efficien

Graph Theoretical Outlier Rejection for 4D Radar Registration in Feature-Poor Environments

ResearchDGX agent

arXiv:2604.14857v1 Announce Type: new Abstract: Automotive 4D imaging radar is well suited for operation in dusty and low-visibility environments, but scan registration remains challenging due to scan

Grok 4.3 can Take in Video and Extract Audio files

IndustryDGX agent

Grok 4.3, an AI model by xAI, has been updated with multimodal capabilities enabling it to process video input and extract audio from video files. This enhancement expands Grok's functionality beyond

Hijacking online reviews: sparse manipulation and behavioral buffering in popularity-biased rating systems

AgentsDGX agent

arXiv:2604.13049v1 Announce Type: cross Abstract: Online reviews and recommendation systems help users navigate overwhelming choice, but they are vulnerable to self-reinforcing distortions. This paper

How Retrieved Context Shapes Internal Representations in RAG

ResearchDGX agent

arXiv:2602.20091v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) enhances large language models (LLMs) by conditioning generation on retrieved external documents, but the effec

How robots learn: A brief, contemporary history

TutorialsDGX agent

Roboticists used to dream big but build small. They’d hope to match or exceed the extraordinary complexity of the human body, and then they’d spend their career refining robotic arms for auto plants.

https://github.com/emollick/tower-of-babel

ApplicationsDGX agent

Tower of Babel is a GitHub repository by researcher Ethan Mollick that likely explores how AI language models handle multilingual tasks and cross-language communication, potentially investigating chal

Integration of Deep Reinforcement Learning and Agent-based Simulation to Explore Strategies Counteracting Information Disorder

AgentsDGX agent

arXiv:2604.13047v1 Announce Type: cross Abstract: In recent years, the spread of fake news has triggered a growing interest in Information Disorders (ID) on social media, a phenomenon that has become

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers

HardwareDGX agent

arXiv:2509.23638v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models face memory and PCIe latency bottlenecks when deployed on commodity hardware. Offloading expert weights to CPU memor

Metric-agnostic Learning-to-Rank via Boosting and Rank Approximation

ApplicationsDGX agent

arXiv:2604.15101v1 Announce Type: cross Abstract: Learning-to-Rank (LTR) is a supervised machine learning approach that constructs models specifically designed to order a set of items or documents bas

ML-based approach to classification and generation of structured light propagation in turbulent media

ResearchDGX agent

arXiv:2604.14208v1 Announce Type: cross Abstract: This work develops machine learning approaches to classify structured light wave beams developing random speckle disturbances as they propagate throug

MLDAS: Machine Learning Dynamic Algorithm Selection for Software-Defined Networking Security

ResearchDGX agent

arXiv:2604.14957v1 Announce Type: cross Abstract: Network security is a critical concern in the digital landscape of today, with users demanding secure browsing experiences and protection of their per

Multi-Modal Manipulation via Multi-Modal Policy Consensus

SafetyDGX agent

arXiv:2509.23468v3 Announce Type: replace-cross Abstract: Effectively integrating diverse sensory modalities is crucial for robotic manipulation. However, the typical approach of feature concatenation

On the Creativity of AI Agents

AgentsDGX agent

arXiv:2604.13242v1 Announce Type: cross Abstract: Large language models (LLMs), particularly when integrated into agentic systems, have demonstrated human- and even superhuman-level performance across

One-shot Compositional 3D Head Avatars with Deformable Hair

ResearchDGX agent

arXiv:2604.14782v1 Announce Type: new Abstract: We propose a compositional method for constructing a complete 3D head avatar from a single image. Prior one-shot holistic approaches frequently fail to

One-shot learning for the complex dynamical behaviors of weakly nonlinear forced oscillators

AgentsDGX agent

arXiv:2604.15181v1 Announce Type: new Abstract: Extrapolative prediction of complex nonlinear dynamics remains a central challenge in engineering. This study proposes a one-shot learning method to ide

Prompt-to-Gesture: Measuring the Capabilities of Image-to-Video Deictic Gesture Generation

ResearchDGX agent

arXiv:2604.14953v1 Announce Type: new Abstract: Gesture recognition research, unlike NLP, continues to face acute data scarcity, with progress constrained by the need for costly human recordings or im

RACER: Retrieval-Augmented Contextual Rapid Speculative Decoding

ResearchDGX agent

arXiv:2604.14885v1 Announce Type: new Abstract: Autoregressive decoding in Large Language Models (LLMs) generates one token per step, causing high inference latency. Speculative decoding (SD) mitigate

Rethinking LLM-Driven Heuristic Design: Generating Efficient and Specialized Solvers via Dynamics-Aware Optimization

ResearchDGX agent

arXiv:2601.20868v2 Announce Type: replace Abstract: Large Language Models (LLMs) have advanced the field of Combinatorial Optimization through automated heuristic generation. Instead of relying on man

Seen-to-Scene: Keep the Seen, Generate the Unseen for Video Outpainting

ResearchDGX agent

arXiv:2604.14648v1 Announce Type: new Abstract: Video outpainting aims to expand the visible content of a video beyond the original frame boundaries while preserving spatial fidelity and temporal cohe

SegWithU: Uncertainty as Perturbation Energy for Single-Forward-Pass Risk-Aware Medical Image Segmentation

ResearchDGX agent

arXiv:2604.15271v1 Announce Type: new Abstract: Reliable uncertainty estimation is critical for medical image segmentation, where automated contours feed downstream quantification and clinical decisio

Sentiment analysis for software engineering: How far can zero-shot learning (ZSL) go?

ResearchDGX agent

arXiv:2604.13826v1 Announce Type: cross Abstract: Sentiment analysis in software engineering focuses on understanding emotions expressed in software artifacts. Previous research highlighted the limita

// Skill Learning for Autonomous Web Agents // Web agents can navigate a page, but ask them to repeat a checkout flow they already completed…

AgentsDGX agent

// Skill Learning for Autonomous Web Agents // Web agents can navigate a page, but ask them to repeat a checkout flow they already completed, and they start from scratch every time. This work introduc

SpaceMind: A Modular and Self-Evolving Embodied Vision-Language Agent Framework for Autonomous On-orbit Servicing

AgentsDGX agent

arXiv:2604.14399v1 Announce Type: new Abstract: Autonomous on-orbit servicing demands embodied agents that perceive through visual sensors, reason about 3D spatial situations, and execute multi-phase

The Code Whisperer: LLM and Graph-Based AI for Smell and Vulnerability Resolution

ResearchDGX agent

arXiv:2604.13114v1 Announce Type: cross Abstract: Code smells and software vulnerabilities both increase maintenance cost, yet they are often handled by separate tools that miss structural context and

'The post all LLMs hate and want you to NOT SEE' ^^ Serious title, you can also discuss this with the LLM, it would be biased against it, as…

AgentsDGX agent

'The post all LLMs hate and want you to NOT SEE' ^^ Serious title, you can also discuss this with the LLM, it would be biased against it, as it's known at this point that memory is the true long term

Thermodynamic Diffusion Inference with Minimal Digital Conditioning

HardwareDGX agent

arXiv:2604.14332v1 Announce Type: new Abstract: Diffusion-model inference and overdamped Langevin dynamics are formally identical. A physical substrate that encodes the score function therefore equili

TOPCELL: Topology Optimization of Standard Cell via LLMs

SafetyDGX agent

arXiv:2604.14237v1 Announce Type: new Abstract: Transistor topology optimization is a critical step in standard cell design, directly dictating diffusion sharing efficiency and downstream routability.

Towards Bridging the Reward-Generation Gap in Direct Alignment Algorithms

SafetyDGX agent

arXiv:2506.09457v3 Announce Type: replace Abstract: Direct Alignment Algorithms (DAAs), such as Direct Preference Optimization (DPO) and Simple Preference Optimization (SimPO), have emerged as efficie

Towards Scalable Lightweight GUI Agents via Multi-role Orchestration

SafetyDGX agent

arXiv:2604.13488v1 Announce Type: new Abstract: Autonomous Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) enable digital automation on end-user devices. Whil

Tug-of-War within A Decade: Conflict Resolution in Vulnerability Analysis via Teacher-Guided Retrieval-Augmented Generations

ResearchDGX agent

arXiv:2604.14172v1 Announce Type: new Abstract: Large Language Models (LLMs) are essential for analyzing and addressing vulnerabilities in cybersecurity. However, among over 200,000 vulnerabilities we

Universal hidden monotonic trend estimation with contrastive learning

ResearchDGX agent

arXiv:2210.09817v3 Announce Type: replace Abstract: In this paper, we describe a universal method for extracting the underlying monotonic trend factor from time series data. We propose an approach rel

v0.21.0-rc1

Local AiDGX agent

Ollama v0.21.0-rc1 is a pre-release that includes improvements to Gemma 4 Tool Calling, adds the latest models to the Ollama App, and fixes issues with launching the OpenClaw TUI. The release enables

XMark: Reliable Multi-Bit Watermarking for LLM-Generated Texts

ResearchDGX agent

arXiv:2604.05242v2 Announce Type: replace Abstract: Multi-bit watermarking has emerged as a promising solution for embedding imperceptible binary messages into Large Language Model (LLM)-generated tex

You gotta pump those numbers up. Those are rookie numbers in this racket.

TutorialsDGX agent

This appears to be a tweet by Jeremy Howard, a prominent figure in AI and machine learning, likely discussing the need to increase certain metrics or performance benchmarks in AI development or a rela

16 Apr 2026

A deep mystery to me is that if I upload writing to a chatbot and ask it for a list of individual improvements, basically everything it give…

TutorialsDGX agent

A deep mystery to me is that if I upload writing to a chatbot and ask it for a list of individual improvements, basically everything it gives me makes the text more punchy and direct and nice to read.

A Function-Centric Perspective on Flat and Sharp Minima

ResearchDGX agent

arXiv:2510.12451v2 Announce Type: replace-cross Abstract: Flat minima are strongly associated with improved generalisation in deep neural networks. However, this connection has proven nuanced in recen

A Lightweight, Transferable, and Self-Adaptive Framework for Intelligent DC Arc-Fault Detection in Photovoltaic Systems

Local AiDGX agent

arXiv:2603.25749v2 Announce Type: replace-cross Abstract: Arc-fault circuit interrupters (AFCIs) are essential for mitigating fire hazards in residential photovoltaic (PV) systems, yet achieving relia

A Review of Diffusion-based Simulation-Based Inference: Foundations and Applications in Non-Ideal Data Scenarios

TutorialsDGX agent

arXiv:2512.23748v2 Announce Type: replace Abstract: For complex simulation problems, inferring parameters often precludes the use of classical likelihood-based techniques due to intractable likelihood

Agentic Conversational Search with Contextualized Reasoning via Reinforcement Learning

AgentsDGX agent

arXiv:2601.13115v2 Announce Type: replace Abstract: Large Language Models (LLMs) have become a popular interface for human-AI interaction, supporting information seeking and task assistance through na

AI for Materials Science starter kit [D]

ResearchDGX agent

This r/MachineLearning discussion post serves as a community-curated beginner's resource for applying artificial intelligence and machine learning to materials science, likely compiling recommended to

b8814

Local AiDGX agent

llama.cpp is a C/C++ library for LLM inference that enables running large language models on consumer hardware. Release b8814 is a specific version in the project's continuous release cycle, which fol

Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data

Local AiDGX agent

arXiv:2604.13688v1 Announce Type: new Abstract: 3D editing refers to the ability to apply local or global modifications to 3D assets. Effective 3D editing requires maintaining semantic consistency by

Calibrated Speculative Decoding: Frequency-Guided Candidate Selection for Efficient Inference

ResearchDGX agent

arXiv:2604.13634v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by letting draft tokens bypass full verification, but conventional frameworks suffer from fre

Confused about plus tier

IndustryDGX agent

This Reddit thread from r/ChatGPT likely reflects common user confusion around what the ChatGPT Plus subscription ($20/month) includes, particularly regarding model access, usage limits, and how it co

Creo: From One-Shot Image Generation to Progressive, Co-Creative Ideation

ResearchDGX agent

arXiv:2604.13956v1 Announce Type: cross Abstract: Text-to-image (T2I) systems enable rapid generation of high-fidelity imagery but are misaligned with how visual ideas develop. T2I systems generate ou

Decentralized Rank Scheduling for Energy-Constrained Multi-Task Federated Fine-Tuning in Edge-Assisted IoV Networks

ApplicationsDGX agent

arXiv:2508.09532v2 Announce Type: replace Abstract: Federated fine-tuning has emerged as a promising approach for adapting foundation models (FMs) to diverse downstream tasks in edge environments. In

Deep Spatially-Regularized and Superpixel-Based Diffusion Learning for Unsupervised Hyperspectral Image Clustering

ResearchDGX agent

arXiv:2604.13307v1 Announce Type: new Abstract: An unsupervised framework for hyperspectral image (HSI) clustering is proposed that incorporates masked deep representation learning with diffusion-base

DiPO: Disentangled Perplexity Policy Optimization for Fine-grained Exploration-Exploitation Trade-Off

SafetyDGX agent

arXiv:2604.13902v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has catalyzed significant advances in the reasoning capabilities of Large Language Models (LLMs).

Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching

SafetyDGX agent

arXiv:2509.21912v2 Announce Type: replace Abstract: Guidance provides a simple and effective framework for posterior sampling by steering the generation process towards the desired distribution. When

← Previous
1…792793794795796…1010
Next →