Loading page…
JAM with a Qwen3.5-4B Researcher scores higher than the evaluated training-free memory systems and larger trained agents across most main benchmark settings; the quoted Memory-R1 variants have higher open-domain LoCoMo F1. On HotpotQA it sustains comparatively strong performance as context grows from 56k to 224k, although its F1 decreases with length. · CiteArk