AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
Agents

feels like we have all the primitives for building powerful autonomous agent model task management tool use/auth memory & context management…

DGX agent

feels like we have all the primitives for building powerful autonomous agent model task management tool use/auth memory & context management self improvement UI/UX what’s missing? verifiers (for self

agentsyohei-nakajima--x
27 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Interpretable Deep Learning for Stock Returns: A Consensus-Bottleneck Asset Pricing Model

DGX agent

arXiv:2512.16251v5 Announce Type: replace-cross Abstract: We introduce the Consensus-Bottleneck Asset Pricing Model (CB-APM), which embeds aggregate analyst consensus as a structural bottleneck, treat

researcharxiv-cs-ai
27 Apr 2026
Agents

Learning Reactive Human Motion Generation from Paired Interaction Data Using Transformer-Based Models

DGX agent

arXiv:2604.22164v1 Announce Type: new Abstract: Recent advances in deep learning have enabled the generation of videos from textual descriptions as well as the prediction of future sequences from inpu

agentsarxiv-cs-cv
27 Apr 2026
Model Releases

Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores

DGX agent

arXiv:2604.22063v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in clinical reasoning and risk assessment. However, their interpretive reliability in critical

model-releasesarxiv-cs-ai
27 Apr 2026
Local Ai

Sunday night: Soju, Seafood Broil with Noodles, and scaffolding my own agents with pi powered by @ollama cloud and @OpenAI models.

DGX agent

This post discusses a developer's weekend project combining soju and seafood broil with noodles while working on building custom AI agents using Ollama's pi-powered infrastructure and OpenAI models. I

local-aiollama--x
27 Apr 2026
Model Releases

UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents

DGX agent

arXiv:2602.07038v2 Announce Type: replace-cross Abstract: Key Information Extraction (KIE) from real-world documents remains challenging due to substantial variations in layout structures, visual qual

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Voice Under Revision: Large Language Models and the Normalization of Personal Narrative

DGX agent

arXiv:2604.22142v1 Announce Type: new Abstract: This study examines how large language model rewriting alters the style and narrative texture of personal narratives. It analyzes 300 personal narrative

researcharxiv-cs-cl
27 Apr 2026
Model Releases

When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models

DGX agent

arXiv:2604.22153v1 Announce Type: cross Abstract: When you ask an AI assistant for advice about your career, your marriage, or a conflict with your family, does it give you the same answer regardless

model-releasesarxiv-cs-ai
27 Apr 2026
Agents

Hermes Agent tip of the day: There are 4 ways to deal with the model while its running, - Message it, by default, it will interrupt the agen…

DGX agent

Hermes Agent tip of the day: There are 4 ways to deal with the model while its running, - Message it, by default, it will interrupt the agent loop, stopping it and making it respond to your new messag

agentsnous-research--x
26 Apr 2026
Model Releases

CaST-POI: Candidate-Conditioned Spatiotemporal Modeling for Next POI Recommendation

DGX agent

arXiv:2604.20845v1 Announce Type: cross Abstract: Next Point-of-Interest (POI) recommendation plays a crucial role in location-based services by predicting users' future mobility patterns. Existing me

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

huggingface: https://huggingface.co/collections/deepseek-ai/deepseek-v4.

DGX agent

DeepSeek-V4 is a collection of models released by DeepSeek-AI on Hugging Face that represents their latest generation of large language models. The collection likely includes various model sizes and c

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

@huggingface On this page: https://huggingface.co/models?other=base_model:quantized:deepseek-ai/DeepSeek-V4-Flash

DGX agent

This post references a Hugging Face Models page filtered to show quantized versions of the DeepSeek-V4-Flash model, a lightweight variant of DeepSeek's V4 language model. The page displays community-q

model-releasessimon-willison--x
24 Apr 2026
Applications

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering

DGX agent

arXiv:2604.21027v1 Announce Type: new Abstract: Electronic health record (EHR) question answering is often handled by LLM-based pipelines that are costly to deploy and do not explicitly leverage the h

applicationsarxiv-cs-ai
24 Apr 2026
Local Ai

Learning Reasoning Reward Models from Expert Demonstration via Inverse Reinforcement Learning

DGX agent

arXiv:2510.01857v3 Announce Type: replace Abstract: Current approaches to improving reasoning in large language models (LLMs) primarily rely on either supervised fine-tuning (SFT) over expert traces o

local-aiarxiv-cs-ai
24 Apr 2026
Model Releases

LLaDA2.0-Uni Released

DGX agent

LLaDA2.0-Uni is a unified diffusion large language model (dLLM) based on Mixture-of-Experts architecture that seamlessly integrates multimodal understanding and generation. The model supports text-to-

model-releasesr-stablediffusion
24 Apr 2026
Model Releases

MathDuels: Evaluating LLMs as Problem Posers and Solvers

DGX agent

arXiv:2604.21916v1 Announce Type: new Abstract: As frontier language models attain near-ceiling performance on static mathematical benchmarks, existing evaluations are increasingly unable to different

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Serialisation Strategy Matters: How FHIR Data Format Affects LLM Medication Reconciliation

DGX agent

arXiv:2604.21076v1 Announce Type: cross Abstract: Medication reconciliation at clinical handoffs is a high-stakes, error-prone process. Large language models are increasingly proposed to assist with t

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

The Feedback Hamiltonian is the Score Function: A Diffusion-Model Framework for Quantum Trajectory Reversal

DGX agent

arXiv:2604.21210v1 Announce Type: cross Abstract: In continuously monitored quantum systems, the feedback protocol of Garcia-Pintos, Liu, and Gorshkov reshapes the arrow of time: a Hamiltonian H_{meas

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

When Agents Look the Same: Quantifying Distillation-Induced Similarity in Tool-Use Behaviors

DGX agent

arXiv:2604.21255v1 Announce Type: new Abstract: Model distillation is a primary driver behind the rapid progress of LLM agents, yet it often leads to behavioral homogenization. Many emerging agents sh

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

A Vision-Language-Action Model for Adaptive Ultrasound-Guided Needle Insertion and Needle Tracking

DGX agent

arXiv:2604.20347v1 Announce Type: cross Abstract: Ultrasound (US)-guided needle insertion is a critical yet challenging procedure due to dynamic imaging conditions and difficulties in needle visualiza

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Accelerating PayPal's Commerce Agent with Speculative Decoding: An Empirical Study on EAGLE3 with Fine-Tuned Nemotron Models

DGX agent

arXiv:2604.19767v1 Announce Type: cross Abstract: We evaluate speculative decoding with EAGLE3 as an inference-time optimization for PayPal's Commerce Agent, powered by a fine-tuned llama3.1-nemotron-

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Adaptive Conformal Anomaly Detection with Time Series Foundation Models for Signal Monitoring

DGX agent

arXiv:2604.20122v1 Announce Type: cross Abstract: We propose a post-hoc adaptive conformal anomaly detection method for monitoring time series that leverages predictions from pre-trained foundation mo

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Available on @ollama ! 🤝🤝

DGX agent

Available on @ollama ! 🤝🤝 Qwen 3.6 27B model is available on Ollama! Use it with all the integrations in Ollama or chat with the model. Chat with the model: ollama run qwen3.6:27b OpenClaw: ollama lau

model-releasesqwen--x
23 Apr 2026
Local Ai

Colorful Talks with Graphs: Human-Interpretable Graph Encodings for Large Language Models

DGX agent

arXiv:2602.10386v2 Announce Type: replace Abstract: Graph problems are fundamentally challenging for large language models (LLMs). While LLMs excel at processing unstructured text, graph tasks require

local-aiarxiv-cs-lg
23 Apr 2026
Safety

Hidden Reliability Risks in Large Language Models: Systematic Identification of Precision-Induced Output Disagreements

DGX agent

arXiv:2604.19790v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed under diverse numerical precision configurations, including standard floating-point formats (e.g.

safetyarxiv-cs-ai
23 Apr 2026
Industry

Imagine every frame of your video, generated live directly from a model. No timeline, no compositor, no render farm. Just exactly what you w…

DGX agent

Imagine every frame of your video, generated live directly from a model. No timeline, no compositor, no render farm. Just exactly what you want to see. Imagine every pixel on your screen, streamed liv

industrycristobal-valenzuela--x
23 Apr 2026
Applications

JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

DGX agent

arXiv:2604.20100v1 Announce Type: new Abstract: Robotic autonomy in open-world environments is fundamentally limited by insufficient data diversity and poor cross-embodiment generalization. Existing r

applicationsarxiv-cs-ro
23 Apr 2026
Agents

Lightweight LLM Agent Memory with Small Language Models

DGX agent

arXiv:2604.07798v3 Announce Type: replace Abstract: Although LLM agents can leverage tools for complex tasks, they still need memory to maintain cross-turn consistency and accumulate reusable informat

agentsarxiv-cs-ai
23 Apr 2026
Local Ai

Picking a model for storytelling support

DGX agent

This discussion likely covers model selection guidance for users creating storytelling visuals with Stable Diffusion, addressing factors like stylistic output and consistency . The r/StableDiffusion c

local-air-stablediffusion
23 Apr 2026
Applications

Quantum Adaptive Self-Attention for Quantum Transformer Models

DGX agent

arXiv:2504.05336v3 Announce Type: replace-cross Abstract: Integrating quantum computing into deep learning architectures is a promising but poorly understood endeavor: when does a quantum layer actual

applicationsarxiv-cs-lg
23 Apr 2026
Industry

Really excellent work by the inference team to serve this model so efficiently! To a significant degree, we have to become an AI inference c…

DGX agent

Sam Altman praises the inference team's work on efficiently serving a model, suggesting that becoming proficient in AI inference is crucial to the field's progress. The post appears to highlight the t

industrysam-altman--x
23 Apr 2026
Research

Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models

DGX agent

arXiv:2508.18609v4 Announce Type: replace-cross Abstract: Post-Training Quantization (PTQ) is a critical strategy for efficient Large Language Models (LLMs) deployment. However, existing scaling laws

researcharxiv-cs-ai
23 Apr 2026
Applications

AutoAdapt: Automated domain adaptation for large language models

DGX agent

Deploying large language models (LLMs) in real-world, high-stakes settings is harder than it should be. In high-stakes settings like law, medicine, and cloud incident response, performance and reliabi

applicationsmicrosoft-research
22 Apr 2026
Research

Beyond One Output: Visualizing and Comparing Distributions of Language Model Generations

DGX agent

arXiv:2604.18724v1 Announce Type: new Abstract: Users typically interact with and evaluate language models via single outputs, but each output is just one sample from a broad distribution of possible

researcharxiv-cs-ai
22 Apr 2026
Research

Cell-Based Representation of Relational Binding in Language Models

DGX agent

arXiv:2604.19052v1 Announce Type: new Abstract: Understanding a discourse requires tracking entities and the relations that hold between them. While Large Language Models (LLMs) perform well on relati

researcharxiv-cs-cl
22 Apr 2026
Safety

Chain-of-Thought as a Lens: Evaluating Structured Reasoning Alignment between Human Preferences and Large Language Models

DGX agent

arXiv:2511.06168v3 Announce Type: replace Abstract: This paper primarily demonstrates a method to quantitatively assess the alignment between multi-step, structured reasoning in large language models

safetyarxiv-cs-ai
22 Apr 2026
Tools

Correction: OpenAI are NOT deprecating that model, the announcement was a mistake

DGX agent

Correction: OpenAI are NOT deprecating that model, the announcement was a mistake Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking in

toolssimon-willison--x
22 Apr 2026
Research

Detecting Data Contamination in Large Language Models

DGX agent

arXiv:2604.19561v1 Announce Type: new Abstract: Large Language Models (LLMs) utilize large amounts of data for their training, some of which may come from copyrighted sources. Membership Inference Att

researcharxiv-cs-ai
22 Apr 2026
Agents

Feasibility of Indoor Frame-Wise Lidar Semantic Segmentation via Distillation from Visual Foundation Model

DGX agent

arXiv:2604.18831v1 Announce Type: new Abstract: Frame-wise semantic segmentation of indoor lidar scans is a fundamental step toward higher-level 3D scene understanding and mapping applications. Howeve

agentsarxiv-cs-cv
22 Apr 2026
Industry

I'm post-training a model with ml-intern. wish me luck!

DGX agent

Clem Delangue, CEO of Hugging Face, shared a post about post-training a model using ml-intern, likely referring to a machine learning internship project or internal tool. The post appears to be a casu

industryclem-delangue--x
22 Apr 2026
Research

Impact of large language models on peer review opinions from a fine-grained perspective: Evidence from top conference proceedings in AI

DGX agent

arXiv:2604.19578v1 Announce Type: cross Abstract: With the rapid advancement of Large Language Models (LLMs), the academic community has faced unprecedented disruptions, particularly in the realm of a

researcharxiv-cs-ai
22 Apr 2026
Research

Lost in the Prompt Order: Revealing the Limitations of Causal Attention in Language Models

DGX agent

arXiv:2601.14152v2 Announce Type: replace-cross Abstract: Large language models exhibit surprising sensitivity to the structure of the prompt, but the mechanisms underlying this sensitivity remain poo

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Owner-Harm: A Missing Threat Model for AI Agent Safety

DGX agent

arXiv:2604.18658v1 Announce Type: cross Abstract: Existing AI agent safety benchmarks focus on generic criminal harm (cybercrime, harassment, weapon synthesis), leaving a systematic blind spot for a d

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

DGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

safetyarxiv-cs-cl
22 Apr 2026
Safety

Reasoning Structure Matters for Safety Alignment of Reasoning Models

DGX agent

arXiv:2604.18946v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve strong performance on complex reasoning tasks but often generate harmful responses to malicious user queries. This

safetyarxiv-cs-ai
22 Apr 2026
Agents

ST-Prune: Training-Free Spatio-Temporal Token Pruning for Vision-Language Models in Autonomous Driving

DGX agent

arXiv:2604.19145v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become central to autonomous driving systems, yet their deployment is severely bottlenecked by the massive computat

agentsarxiv-cs-ai
22 Apr 2026
Industry

Team is hard at work together with @steipete to make OpenAI models and ecosystem be the obvious way to to enjoy your claw. A lot more to com…

DGX agent

Team is hard at work together with @steipete to make OpenAI models and ecosystem be the obvious way to to enjoy your claw. A lot more to come next week, but a reminder that you can use OpenClaw as par

industrysam-altman--x
22 Apr 2026
Tools

This is why I won't use proprietary hosted embedding models myself - I am more than happy to pay for a hosted solution (cheaper, faster and …

DGX agent

This is why I won't use proprietary hosted embedding models myself - I am more than happy to pay for a hosted solution (cheaper, faster and more convenient than self-hosting) but I want an open weight

toolssimon-willison--x
22 Apr 2026
← Previous
1…207208209210211…1271
Next →