AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations

DGX agent

arXiv:2601.03779v2 Announce Type: replace Abstract: We explore intrinsic dimension (ID) of LLM representations as a marker of linguistic complexity. Specifically, we test whether ID differences across

researcharxiv-cs-cl
27 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

UR^2: Unify RAG and Reasoning through Reinforcement Learning

DGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threa…

DGX agent

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threads. I don’t think we’ve cracked the right UI for managing ag

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate'…

DGX agent

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST t…

DGX agent

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST time I feel I have a frontier model running on my computer. T

model-releasesclem-delangue--x
26 Apr 2026
Industry

5.5 is so earnest 'little engine that could' energy

DGX agent

Sam Altman praised OpenAI's o1 model (version 5.5) for its earnest, persistent approach to problem-solving, comparing it to the 'little engine that could' mentality. The comment reflects Altman's pers

industrysam-altman--x
25 Apr 2026
Model Releases

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹C…

DGX agent

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹Claude Code: Set model to deepseek-v4-pro[1m] to unlock 1M conte

model-releasesdeepseek--x
25 Apr 2026
Research

http://reddit.com/r/LocalLLaMA

DGX agent

r/LocalLLaMA is a subreddit community dedicated to discussing and sharing resources about running large language models locally on personal computers, covering topics like model optimization, hardware

researchnous-research--x
25 Apr 2026
Model Releases

WHY ARE YOU LIKE THIS

DGX agent

@scottjla on Twitter in reply to my pelican riding a bicycle benchmark: I feel like we need to stack these tests now I checked to confirm that the model (ChatGPT Images 2.0) added the 'WHY ARE YOU LIK

model-releasessimon-willison
25 Apr 2026
Model Releases

4⃣4⃣4⃣4⃣

DGX agent

4⃣4⃣4⃣4⃣ Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit

model-releasestogether-ai--x
24 Apr 2026
Model Releases

A-THENA: Early Intrusion Detection for IoT with Time-Aware Hybrid Encoding and Network-Specific Augmentation

DGX agent

arXiv:2604.21623v1 Announce Type: cross Abstract: The proliferation of Internet of Things (IoT) devices has significantly expanded attack surfaces, making IoT ecosystems particularly susceptible to so

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Absorber LLM: Harnessing Causal Synchronization for Test-Time Training

DGX agent

arXiv:2604.20915v1 Announce Type: cross Abstract: Transformers suffer from a high computational cost that grows with sequence length for self-attention, making inference in long streams prohibited by

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security

DGX agent

arXiv:2601.18491v2 Announce Type: replace Abstract: The rise of AI agents introduces complex safety and security challenges arising from autonomous tool use and environmental interactions. Current gua

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents

DGX agent

arXiv:2604.21134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently misread values, hallucinate details, and confuse overlapping elements in charts. Current approaches rely solely

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

BioMiner: A Multi-modal System for Automated Mining of Protein-Ligand Bioactivity Data from Literature

DGX agent

arXiv:2604.21508v1 Announce Type: new Abstract: Protein-ligand bioactivity data published in the literature are essential for drug discovery, yet manual curation struggles to keep pace with rapidly gr

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion

DGX agent

arXiv:2604.04395v2 Announce Type: replace Abstract: 3D conducting motion generation aims to synthesize fine-grained conductor motions from music, with broad potential in music education, virtual perfo

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

DGX agent

arXiv:2604.21308v1 Announce Type: cross Abstract: Enterprise LLM agents can dramatically improve workplace productivity, but their core capability, retrieving and using internal context to act on a us

model-releasesarxiv-cs-cl
24 Apr 2026
Research

climt-paraformer: Stable Emulation of Convective Parameterization using a Temporal Memory-aware Transformer

DGX agent

arXiv:2604.21085v1 Announce Type: cross Abstract: Accurate representation of moist convective sub-grid-scale processes remains a major challenge in global climate models, as traditional parameterizati

researcharxiv-cs-lg
24 Apr 2026
Model Releases

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors

DGX agent

arXiv:2604.21241v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) models often use intermediate representations to connect multimodal inputs with continuous control, yet spatial guidanc

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Data-Driven Open-Loop Simulation for Digital-Twin Operator Decision Support in Wastewater Treatment

DGX agent

arXiv:2604.20935v1 Announce Type: cross Abstract: Wastewater treatment plants (WWTPs) need digital-twin-style decision support tools that can simulate plant response under prescribed control plans, to

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

DeepSeek v4 just dropped

DGX agent

DeepSeek has released v4, its latest model iteration. The announcement was made by Clem Delangue on X (formerly Twitter). This likely represents a significant update to DeepSeek's AI capabilities, tho

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

DeepSeek V4 Pro is now available on Together AI. DeepSeek V4 Flash coming soon. Try it now: http://www.together.ai/models/deepseek-v4-pro#

DGX agent

DeepSeek V4 Pro is now available through Together AI's model platform, with the faster DeepSeek V4 Flash variant expected to launch soon. Together AI is offering users the ability to access and test D

model-releasestogether-ai--x
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

DGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Fine-Tuning Regimes Define Distinct Continual Learning Problems

DGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc

model-releasesarxiv-cs-lg
24 Apr 2026
Safety

Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling

DGX agent

arXiv:2506.09998v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can often accurately describe probability distributions using natural language, yet they still struggle to genera

safetyarxiv-cs-cl
24 Apr 2026
Research

Generative Discovery of Magnetic Insulators under Competing Physical Constraints

DGX agent

arXiv:2604.21073v1 Announce Type: cross Abstract: Discovering materials that must simultaneously satisfy multiple competing constraints remains a central challenge in computational materials design, p

researcharxiv-cs-ai
24 Apr 2026
Model Releases

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

DGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

DGX agent

arXiv:2604.21901v1 Announce Type: cross Abstract: As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopt

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GPT-5.5 now available in Deep Agents!

DGX agent

GPT-5.5 now available in Deep Agents! GPT-5.5 is now available in the API. The model brings higher intelligence and stronger token efficiency to complex work, helping tasks get done with fewer retries

model-releasesharrison-chase--x
24 Apr 2026
Safety

HARBOR: Automated Harness Optimization

DGX agent

arXiv:2604.20938v1 Announce Type: cross Abstract: Long-horizon language-model agents are dominated, in lines of code and in operational complexity, not by their underlying model but by the harness tha

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Here's DeepSeek v4 Pro. Added to the playable gallery as well.

DGX agent

Here's DeepSeek v4 Pro. Added to the playable gallery as well. Media I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 30

model-releasesethan-mollick--x
24 Apr 2026
Tutorials

Information Bottleneck-Guided Heterogeneous Graph Learning for Interpretable Neurodevelopmental Disorder Diagnosis

DGX agent

arXiv:2502.20769v3 Announce Type: replace Abstract: Developing interpretable models for neurodevelopmental disorders (NDDs) diagnosis presents significant challenges in effectively encoding, decoding,

tutorialsarxiv-cs-cv
24 Apr 2026
Model Releases

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

DGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

model-releasesarxiv-cs-cv
24 Apr 2026
Research

Job Skill Extraction via LLM-Centric Multi-Module Framework

DGX agent

arXiv:2604.21525v1 Announce Type: new Abstract: Span-level skill extraction from job advertisements underpins candidate-job matching and labor-market analytics, yet generative large language models (L

researcharxiv-cs-cl
24 Apr 2026
Applications

LAF-Based Evaluation and UTTL-Based Learning Strategies with MIATTs

DGX agent

arXiv:2604.20944v1 Announce Type: cross Abstract: In many real-world machine learning (ML) applications, the true target cannot be precisely defined due to ambiguity or subjectivity information. To ad

applicationsarxiv-cs-ai
24 Apr 2026
Research

LatRef-Diff: Latent and Reference-Guided Diffusion for Facial Attribute Editing and Style Manipulation

DGX agent

arXiv:2604.21279v1 Announce Type: new Abstract: Facial attribute editing and style manipulation are crucial for applications like virtual avatars and photo editing. However, achieving precise control

researcharxiv-cs-cv
24 Apr 2026
Research

Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training

DGX agent

arXiv:2604.21265v1 Announce Type: new Abstract: We show that pre-training a Transformer on music before language significantly accelerates language acquisition. Using piano performances (MAESTRO datas

researcharxiv-cs-cl
24 Apr 2026
Safety

Logic Jailbreak: Efficiently Unlocking LLM Safety Restrictions Through Formal Logical Expression

DGX agent

arXiv:2505.13527v3 Announce Type: replace-cross Abstract: Despite substantial advancements in aligning large language models (LLMs) with human values, current safety mechanisms remain susceptible to j

safetyarxiv-cs-ai
24 Apr 2026
Research

Losing our Tail, Again: (Un)Natural Selection & Multilingual LLMs

DGX agent

arXiv:2507.03933v3 Announce Type: replace Abstract: Multilingual Large Language Models considerably changed how technologies influence language. While previous technologies could mediate or assist hum

researcharxiv-cs-cl
24 Apr 2026
Model Releases

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

DGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

model-releasesarxiv-cs-cv
24 Apr 2026
Safety

Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

DGX agent

arXiv:2604.21092v1 Announce Type: new Abstract: Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processe

safetyarxiv-cs-ai
24 Apr 2026
Tutorials

Mixture of Sequence: Theme-Aware Mixture-of-Experts for Long-Sequence Recommendation

DGX agent

arXiv:2604.20858v1 Announce Type: cross Abstract: Sequential recommendation has rapidly advanced in click-through rate prediction due to its ability to model dynamic user interests. A key challenge, h

tutorialsarxiv-cs-ai
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

DGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

model-releasesarxiv-cs-cv
24 Apr 2026
Applications

Optimizing Diffusion Priors with a Single Observation

DGX agent

arXiv:2604.21066v1 Announce Type: new Abstract: While diffusion priors generate high-quality posterior samples across many inverse problems, they are often trained on limited training sets or purely s

applicationsarxiv-cs-cv
24 Apr 2026
Model Releases

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Nex…

DGX agent

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Next event, including our eighth generation TPUs (TPUt for infe

model-releasesgoogle-ai--x
24 Apr 2026
← Previous
1…550551552553554…1371
Next →