AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,598 results
Tools

This is a really big deal - it's easy to run into nasty bills with Cloud Run if your site attracts aggressive scrapers, spending caps make i…

DGX agent

This is a really big deal - it's easy to run into nasty bills with Cloud Run if your site attracts aggressive scrapers, spending caps make it a whole lot safer to run small projects on Anouncing Spend

toolssimon-willison--x
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

DGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

safetyarxiv-cs-ai
22 Apr 2026
Research

Amortized Inverse Kinematics via Graph Attention for Real-Time Human Avatar Animation

DGX agent

arXiv:2604.16629v1 Announce Type: new Abstract: Inverse kinematics (IK) is a core operation in animation, robotics, and biomechanics: given Cartesian constraints, recover joint rotations under a known

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods

DGX agent

arXiv:2510.07143v3 Announce Type: replace Abstract: Recent efforts to accelerate inference in Multimodal Large Language Models (MLLMs) have largely focused on visual token compression. The effectivene

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation

DGX agent

arXiv:2604.18468v1 Announce Type: new Abstract: Closed-loop simulation is a core component of autonomous vehicle (AV) development, enabling scalable testing, training, and safety validation before rea

safetyarxiv-cs-cv
21 Apr 2026
Safety

Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization

DGX agent

arXiv:2604.17188v1 Announce Type: new Abstract: Multi-role dialogue summarization requires modeling complex interactions among multiple speakers while preserving role-specific information and factual

safetyarxiv-cs-cl
21 Apr 2026
Research

Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation

DGX agent

arXiv:2604.17020v1 Announce Type: new Abstract: Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling

DGX agent

arXiv:2604.17794v1 Announce Type: new Abstract: The democratization of ubiquitous AI hinges on deploying sophisticated reasoning capabilities on resource-constrained devices. However, Small Language M

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China

DGX agent

arXiv:2510.08986v2 Announce Type: replace Abstract: We introduce CAPC-CG, the Chinese Adaptive Policy Communication (Central Government) Corpus, the first open dataset of Chinese policy directives ann

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations

DGX agent

arXiv:2507.20409v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting helps models think step by step. But naive CoT breaks down in visually grounded social tasks, where models must per

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Culture-Aware Humorous Captioning: Multimodal Humor Generation across Cultural Contexts

DGX agent

arXiv:2604.18091v1 Announce Type: new Abstract: Recent multimodal large language models have shown promising ability in generating humorous captions for images, yet they still lack stable control over

safetyarxiv-cs-cl
21 Apr 2026
Research

DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies

DGX agent

arXiv:2503.14324v3 Announce Type: replace-cross Abstract: The differing representation spaces required for visual understanding and generation pose a challenge in unifying them within the autoregressi

researcharxiv-cs-cl
21 Apr 2026
Safety

Dynamic Emotion and Personality Profiling for Multimodal Deception Detection

DGX agent

arXiv:2604.17037v1 Announce Type: new Abstract: Deception detection is of great significance for ensuring information security and conducting public opinion analysis, with personality factors and emot

safetyarxiv-cs-cl
21 Apr 2026
Research

Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation

DGX agent

arXiv:2502.13637v2 Announce Type: replace Abstract: Human affordance learning investigates contextually relevant novel pose prediction such that the estimated pose represents a valid human action with

researcharxiv-cs-cv
21 Apr 2026
Safety

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

DGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

safetyarxiv-cs-cl
21 Apr 2026
Safety

Human-Centered Supervision for Sentiment Analysis in Telugu: A Systematic Inquiry Beyond Accuracy

DGX agent

arXiv:2508.01486v3 Announce Type: replace Abstract: Sentiment analysis for low-resource languages remains challenging in an era where interpretability, human alignment, and fairness are increasingly n

safetyarxiv-cs-cl
21 Apr 2026
Applications

I built a web analytics system to track traffic and events across all my @Replit projects in one place. Drop-in @posthog stack: → Full analy…

DGX agent

I built a web analytics system to track traffic and events across all my @Replit projects in one place. Drop-in @posthog stack: → Full analytics dashboard → Reverse proxy (beat ad blockers) → AI skill

applicationsreplit--x
21 Apr 2026
Model Releases

I find that open weights models over-perform on benchmarks compared to actual real-world usage, and Kimi feels like no exception. For exampl…

DGX agent

I find that open weights models over-perform on benchmarks compared to actual real-world usage, and Kimi feels like no exception. For example, a small amount of use will show that Kimi is not as good

model-releasesethan-mollick--x
21 Apr 2026
Model Releases

Jupiter-N Technical Report

DGX agent

arXiv:2604.17429v1 Announce Type: new Abstract: We present Jupiter-N, a hybrid reasoning model post-trained from Nemotron 3 Super, a fully open-source 120 billion parameter LLM. We target three object

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

K2.6 + hermes = 4 hr session setting up qwen 3.6 training regime on dgx spark with the current autnomous session currently lasting 70+ min w…

DGX agent

K2.6 + hermes = 4 hr session setting up qwen 3.6 training regime on dgx spark with the current autnomous session currently lasting 70+ min without any prompting. Kimi with hermes is next level. @NousR

model-releasesnous-research--x
21 Apr 2026
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Learning-Based Sparsification of Dynamic Graphs in Robotic Exploration Algorithms

DGX agent

arXiv:2604.16509v1 Announce Type: cross Abstract: Many robotic exploration algorithms rely on graph structures for frontier-based exploration and dynamic path planning. However, these graphs grow rapi

safetyarxiv-cs-lg
21 Apr 2026
Safety

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

DGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

safetyarxiv-cs-cl
21 Apr 2026
Safety

Modeling User Exploration Saturation: When Recommender Systems Should Stop Pushing Novelty

DGX agent

arXiv:2604.16419v1 Announce Type: cross Abstract: Fairness-aware recommender systems often mitigate bias by increasing exposure to under-represented or long-tail content, commonly through mechanisms t

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines

DGX agent

arXiv:2604.17340v1 Announce Type: new Abstract: Clinical guidelines, typically developed by independent specialty societies, inherently exhibit substantial fragmentation, redundancy, and logical contr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions

DGX agent

arXiv:2604.16493v1 Announce Type: cross Abstract: Natural Language to SQL (NL2SQL) technology empowers non-expert users to query relational databases without requiring SQL expertise. While large langu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

DGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Plasticity Loss in Deep Reinforcement Learning: A Survey

DGX agent

arXiv:2411.04832v3 Announce Type: replace-cross Abstract: Plasticity refers to a network's ability to adapt to changing data distributions, which is crucial for the successful training of deep reinfor

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?

DGX agent

arXiv:2604.17338v1 Announce Type: cross Abstract: Unlike code completion, debugging requires localizing faults and applying targeted edits. We observe that frontier LLMs often regenerate correct but o

model-releasesarxiv-cs-cl
21 Apr 2026
Research

PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues

DGX agent

arXiv:2604.18354v1 Announce Type: new Abstract: Emotion plays a pivotal role in shaping negotiation outcomes, influencing trust, cooperation, and long-term relationships. Developing negotiation dialog

researcharxiv-cs-cl
21 Apr 2026
Safety

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

DGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

safetyarxiv-cs-cv
21 Apr 2026
Research

REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations

DGX agent

arXiv:2604.17289v1 Announce Type: new Abstract: Supervised fine-tuning of large language models relies on human-annotated data, yet annotation pipelines routinely involve multiple crowdworkers of hete

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Scaling Human-AI Coding Collaboration Requires a Governable Consensus Layer

DGX agent

arXiv:2604.17883v1 Announce Type: cross Abstract: Vibe coding produces correct, executable code at speed, but leaves no record of the structural commitments, dependencies, or evidence behind it. Revie

model-releasesarxiv-cs-lg
21 Apr 2026
Research

ScenarioControl: Vision-Language Controllable Vectorized Latent Scenario Generation

DGX agent

arXiv:2604.17147v1 Announce Type: new Abstract: We introduce ScenarioControl, the first vision-language control mechanism for learned driving scenario generation. Given a text prompt or an input image

researcharxiv-cs-cv
21 Apr 2026
Research

SentiAvatar: Towards Expressive and Interactive Digital Humans

DGX agent

arXiv:2604.02908v2 Announce Type: replace Abstract: We present SentiAvatar, a framework for building expressive interactive 3D digital humans, and use it to create SuSu, a virtual character that speak

researcharxiv-cs-cv
21 Apr 2026
Industry

SpaceX says it's working with Cursor to build 'the world's most useful models' and it has the right to acquire Cursor for 60B or pay 10B for the partnership (New York Times)

DGX agent

New York Times: SpaceX says it's working with Cursor to build “the world's most useful models” and it has the right to acquire Cursor for 60B or pay 10B for the partnership — The potential acquisition

industrytechmeme
21 Apr 2026
Local Ai

ST-pi: Structured SpatioTemporal VLA for Robotic Manipulation

DGX agent

arXiv:2604.17880v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have achieved great success on general robotic tasks, but still face challenges in fine-grained spatiotemporal man

local-aiarxiv-cs-cv
21 Apr 2026
Research

Stable Language Guidance for Vision-Language-Action Models

DGX agent

arXiv:2601.04052v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive capabilities in generalized robotic control; however, they remain notoriously

researcharxiv-cs-cl
21 Apr 2026
Research

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning

DGX agent

arXiv:2604.17551v1 Announce Type: new Abstract: Standard approaches to goal-conditioned reinforcement learning (GCRL) that rely on temporal-difference learning can be unstable and sample-inefficient d

researcharxiv-cs-lg
21 Apr 2026
Applications

Syenta raises $26M in funding to speed up chip interconnect production

DGX agent

Australian chip manufacturing startup Syenta Inc. today announced that it has raised 26 million in funding to expand its production capacity. Playground Global and Australia’s National Reconstruction

applicationssiliconangle
21 Apr 2026
Safety

Synthia: Scalable Grounded Persona Generation from Social Media Data

DGX agent

arXiv:2507.14922v2 Announce Type: replace Abstract: Persona-driven simulations are increasingly used in computational social science, yet their validity critically depends on the fidelity of the under

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

TeleEmbedBench: A Multi-Corpus Embedding Benchmark for RAG in Telecommunications

DGX agent

arXiv:2604.17778v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in the telecommunications domain for critical tasks, relying heavily on Retrieval-Augmented Gener

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models

DGX agent

arXiv:2604.18107v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) achieve remarkable performance in sequential decision-making but remain fragile to subtle environmental shifts, suc

researcharxiv-cs-cv
21 Apr 2026
Tools

The next Slack won't look like Slack

DGX agent

This post likely discusses how the next generation of team communication and collaboration platforms will diverge from Slack's current design paradigm and feature set, possibly exploring new UI/UX app

toolsswyx--x
21 Apr 2026
Research

Towards Disentangled Preference Optimization Dynamics Beyond Likelihood Displacement

DGX agent

arXiv:2604.18239v1 Announce Type: new Abstract: Preference optimization is widely used to align large language models (LLMs) with human preferences. However, many margin-based objectives suppress the

researcharxiv-cs-lg
21 Apr 2026
Model Releases

TSVer: A Benchmark for Fact Verification Against Time-Series Evidence

DGX agent

arXiv:2511.01101v2 Announce Type: replace Abstract: Reasoning over temporal and numerical data, such as time series, is a crucial aspect of fact-checking. While many systems have recently been develop

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Untrained CNNs Match Backpropagation at V1: A Systematic RSA Comparison of Four Learning Rules Against Human fMRI

DGX agent

arXiv:2604.16875v1 Announce Type: new Abstract: A central question in computational neuroscience is whether the learning rule used to train a neural network determines how well its internal representa

safetyarxiv-cs-lg
21 Apr 2026
Research

Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models

DGX agent

arXiv:2506.00065v2 Announce Type: replace Abstract: Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly und

researcharxiv-cs-cl
21 Apr 2026
← Previous
1…357358359360361…367
Next →