Hacker News
Ox-Alpha Is GLM?
swiftcoder
|next
[-]
Its amazing to me that providers haven't added any sort of masking of the prompt in the thinking traces to avoid prompt extraction via this sort of trivial attack
gvkhna
|next
|previous
[-]
ggcr
|root
|parent
|next
[-]
e9
|root
|parent
|previous
[-]
gvkhna
|root
|parent
|next
[-]
It feels like glm flash, and there was a report zhipu had secured a huge new cluster suggesting they have the capacity. My guess anyway.
https://www.tomshardware.com/tech-industry/artificial-intell...
jerrythegerbil
|next
|previous
[-]
But while we’re “guessing”: Xiaomi MiMO
tadkar
|next
|previous
[-]
mogili
|next
|previous
[-]
gvkhna
|root
|parent
|next
[-]
volf_
|next
|previous
[-]
My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot.
MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant than the lab behind ox-alpha, so less likely).
nylonstrung
|root
|parent
|next
[-]
minimaxir
|root
|parent
|next
|previous
[-]
Bolwin
|root
|parent
|next
|previous
[-]
The only question now is if it's 5.3v, 5.4/5.5 or a dedicated flash/vision model
petesergeant
|next
|previous
[-]
stingraycharles
|root
|parent
[-]
Tepix
|root
|parent
|next
[-]
petesergeant
|root
|parent
|previous
[-]
behnamoh
|previous
[-]
walrus01
|root
|parent
|next
[-]
minimaxir
|root
|parent
|next
|previous
[-]
Given the traction the model has received, it is extremely newsworthy to know who's developing and hosting it.