# mode-io/vllm-jev > A vLLM-based server for Jev decision models: given a question and candidate answers, it returns a label and probability for each candidate. Runs on Linux with NVIDIA GPUs and on Apple Silicon via MLX/MPS. - Magnitude: 2.6 out of 10 — Early signal - Stars: 100 total · +21 stars today, ≈ 32 by evening - Star trust: star growth looks organic - Category: Language models · Language: Python · License: Apache-2.0 · Created: 2026-09-24 · Last push: 2026-09-29 - GitHub: https://github.com/mode-io/vllm-jev · Page: https://gitnova.dev/en/r/mode-io/vllm-jev ## Useful for - Deploy an HTTP server to classify customer intent across candidate labels - Run local Valen inference on a Mac for image-based questions - Handle 40 concurrent text Choice requests on a single A800 ## Why it’s here - 21 stars so far today, about 32 expected by the end of the day. - The spike has held for 3 days in a row — not a one-off blip. - The repository is 5 days old and already has 100 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #158. ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 2 - Issues and pull requests: 1 - Watchers: 1 - Average over the last week: 19 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 4 ## Stars per day, last 10 days (oldest → newest, today is partial) 2026-09-20 … 2026-09-29: 0, 0, 0, 0, 10, 29, 13, 11, 18, 21 ## Spotted in now - Top new repositories this week: #158 ## Similar by description 1. **firelex/jeff** — 8.2 · Early signal · Language models · Python · +282 stars today, ≈ 491 by evening Small fine-tuned Qwen3.5 and Gemma 4 models for zero-shot classification: given a situation description and a list of options, they return a calibrated probability for each option in a single forward pass. Full card: https://gitnova.dev/en/r/firelex/jeff.md 2. **Liuziyu77/Valen** — 4.3 · Early signal · Language models · Python · +22 stars today, ≈ 47 by evening A multimodal decision model built on a Qwen3.5-0.8B/2B backbone: it takes text, images and video with an instruction and returns probabilities over supplied candidates without generating answer tokens. The repo includes model code, data… Full card: https://gitnova.dev/en/r/Liuziyu77/Valen.md 3. **ekzhang/openjev-sglang** — 0.6 · Cooling · Language models · Python · +0 stars today, ≈ 0 by evening HTTP server implementing the TypeSafe/Jev API for structured text classification on Qwen3.6-35B-A3B via SGLang; returns answer probabilities without generating a chain of thought. Full card: https://gitnova.dev/en/r/ekzhang/openjev-sglang.md 4. **bespokelabsai/nimble** — 1.8 · Cooling · Language models · Python · +8 stars today, ≈ 14 by evening Nimble is a model and training recipe for fast typed decisions over text: given a flat schema of enum and boolean fields, it picks an answer and returns probabilities for each option. It targets routing, condition checks, and policy… Full card: https://gitnova.dev/en/r/bespokelabsai/nimble.md 5. **zhulinchng/jevper** — 0.6 · Steady · Language models · Python · +0 stars today, ≈ 0 by evening A classification wrapper over OpenAI-compatible clients: instead of prose it returns typed questions (noul, choice, score) with probabilities and confidence. Works with any client exposing responses.create or chat.completions.create,… Full card: https://gitnova.dev/en/r/zhulinchng/jevper.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-09-29 14:30 UTC, updated every 30 minutes.