Add Vast.ai Mem0 Qwen3.5 launcher - #3
Draft
jinuk0211 wants to merge 2 commits into
Draft
Conversation
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
What changed
Qwen/Qwen3.5-4BWhy
The existing Mem0 preset assumes already-running local endpoints and does not provide a reproducible Vast.ai bootstrap path. Qwen3.5 also requires a newer serving stack than the repository-wide pinned Transformers dependency, so the launcher keeps vLLM setup separate from the broad all-method requirements.
User impact
Users can clone this branch on a CUDA-enabled Vast.ai instance and start a one-query LongMemEval smoke run with a single script invocation.
MAX_QUERIES=0runs every query in the configured five-sample preset, whileMAX_TEST_SAMPLEScan expand that sample limit.Validation
bash -npip install --dry-run --ignore-installed -r requirements-mem0-vast.txtresolves successfullygit diff --cached --checkpassesGPU model loading was not executed on the Windows development host because no local CUDA/WSL runtime is available.