No datacenter needed now, just run it on your computer. Anything with 12gb+ of vram can run it comfortably at low quants. If they can come out with a good MOE model it will crush the competition. At this current rate of progress, China will out-compete foreign AI development within a year or two.

  • lurkerlady [she/her]@hexbear.netOP
    link
    fedilink
    English
    arrow-up
    7
    ·
    2 days ago

    opencode has a hosting org attached to it that provides free preconfigured llm runtime to anyone with opencode installed. it can be a way to access deepseek v4 flash without having good computer specs. but its got a pretty long queue during the day when everyone wants to use it.