proomt

Search

Search posts, papers, and topics

All posts

Hugging Face Daily PapersHaoyang Su, Weiran Huang1 min readpaperadvanced

JevSpawn: Adaptive Agentic Inference through Compositional Action Spaces

Summary

JevSpawn introduces a compositional policy that maps natural‑language task specs to finite‑field probabilistic action spaces, enabling parallel action spawning and feedback‑driven branch selection. Benchmarks show it outperforms several LLM agent baselines while reducing inference latency.

  • LLM agents generate actions token‑by‑token, which is slow; JevSpawn replaces this with fast finite‑field predictions.
  • The method composes natural‑language specifications into a finite action space without pre‑defining fields.
  • Parallel spawning of actions plus feedback‑driven branch selection cuts repeated generation and context computation.
  • Evaluated on eight tasks against seven baselines and a TypeSafe Jev variant, showing higher success rates and faster navigation.

Engineers building LLM‑driven agents need faster, structured inference to handle complex tasks efficiently.

6/10

Related reading

  1. CodeMidas: Scaling Agentic Coding RL Environments from Code Itself

    CodeMidas builds RL environments directly from open‑source code: agents explore a repo, infer a spec, generate tests from the original implementation, and filter tasks via execution checks. The pipeline yields 5,545 high‑quality coding tasks across 23 languages and 15 domains. Training the MiMo‑V2.5 agent with GRPO on this dataset improves benchmark scores by 8‑18% (e.g., DeepSWE +11.7%, ProgramB…

    Hugging Face Daily Papersarxiv.org1 minpaper
  2. HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness

    HarnessVLN introduces a zero‑shot, training‑free embodied navigation framework that wraps a multimodal LLM in an "Agent Harness" – a tool‑based protocol that validates planner actions against spatial evidence, tracks progress with hierarchical event memory, and maintains a persistent spatiotemporal graph for recovery. The system works for instruction‑following and object‑goal tasks, achieving 60.…

    Hugging Face Daily Papersarxiv.org1 minpaper