Strands Decider 2B: a small, open-source, decision model

(strandsagents.com)

123 points | by gmays 4 hours ago

13 comments

  • adenta 3 hours ago
    At this point I can't wait for a comedian to release a decision model backed by humans.

    Meet Jerry- it's literally a guy named Jerry answering your questions.

  • mattvr 51 minutes ago
    Why is everyone calling binary choices `noul`? Does this have some meaning or is it just copying Jev’s API?
  • keyle 3 hours ago
    Fantastically well written. It's rare for me to be able to understand what the AI gurus are talking about, and this was written by humans for humans.

    It can technically be used for a lot of use cases, I'd like people to chime in on ideas on this?

    • nryoo 2 hours ago
      Picking lunch menu..?
  • hrpnk 26 minutes ago
    clef from cloudflare runs on llama.cpp - being locked-in to strands cli would be a bummer and will slow down adoption.

    Since it's a LoRa on Qwen, I assume this is runnable via llama.cpp. Pity that the PEFT/LoRa->GGUF translation is left to the user. Anyone got past:

        $ uv run --with transformers==5.19.0 convert_lora_to_gguf.py ~/Downloads/lora --dry-run --verbose
        [...]
          File "/Users/user/repos/llama.cpp/conversion/base.py", line 630, in map_tensor_name
        raise ValueError(f"Can not map tensor {name!r}")
        ValueError: Can not map tensor 'layers.0.linear_attn.in_proj_a.weight'
  • SubiculumCode 1 hour ago
    Are any of these multimodal yet? I'd love to try asking a model with calibrated probabilities to answer question like, "do these shapes match?". Sure, you can ask a LLM....
  • soltanov 1 hour ago
    Benchmark calibration does not establish reliability on unfamiliar production inputs.
  • miguelspizza 4 hours ago
    This is a great model. I've been running it on device in chrome extension to filter things like email.

    It is just the right mix of size, capability and speed to make it generally useful for adhoc bulk classification tasks.

    For those wanting to run it in browser: https://huggingface.co/alxnahas/strands-decider-2B-webgpu

  • teruakohatu 2 hours ago
    Any idea how well this would run on a CPU?
    • gopalv 1 hour ago
      On my M3 mac, it works okay inside a docker container with just CPU.

      { "model": "strands-decider-2B-hobson-v19", "answers": { "is_urgent": { "type": "noul", "noul": 0.8287 } }, "usage": { "input_tokens": 86, "output_tokens": 1 }, "latency_ms": 1732.17 }

      This is how I got it running - https://gist.github.com/2891eb0db9ea92c1a4e860d44f556292

      There's a lot more to be done if we optimize for MLX & let it run on a Mac mini instead of the docker wrapper.

    • avereveard 1 hour ago
      About half a second per decision on six cores
  • davvie 2 hours ago
    Looks really nice, I think I could use it on my Mac mini for some smaller automations
  • stephantul 54 minutes ago
    2B being called small is such a sign of the times
  • yieldcrv 1 hour ago
    a strand type game
  • lin7c 25 minutes ago
    [flagged]
  • Fluid_Mechanics 3 hours ago
    [flagged]