logoalt Hacker News

Ox Alpha

71 pointsby mtokmak06yesterday at 11:56 PM59 commentsview on HN

https://twitter.com/OpenRouter/status/2090544970923184269, https://xcancel.com/OpenRouter/status/2090544970923184269


Comments

fedposttoday at 1:54 AM

It's Chinese. Won't answer anything about Tiananmen Square but will gleefully give you instructions to perform various electronic warfare attacks that opus and fable instantly refuse.

show 2 replies
walrus01today at 1:35 AM

I highly recommend feeding all your proprietary data and confidential personal information into this model as quickly as possible. What could possibly go wrong?!

In terms of equivalence of suspicion, this is the external inference provider equivalent of getting free steak that was smuggled out of a grocery store inside somebody's pants.

thih9today at 5:19 AM

> It is free.

> This time, the provider does not train on your prompts or completions.

Interesting. And offering product at cost seems exactly the move that a US VC company would make. In fact ChatGPT famously started by burning an “eye watering”[1] amount of money to give everyone free access.

[1]: https://xcancel.com/sama/status/1599669571795185665?lang=en

minimaxirtoday at 5:11 AM

Model is suspiciously fast and has a low reported output token count (using via OpenRouter's Chat), both of which aren't representative of models from the big Chinese labs. Odd.

AnodicElegytoday at 1:26 AM

"Prompts and completions are retained by the provider and are not used for training..."

I'm curious what the model provider is using the prompt/response pairs for, in that case. They aren't offering a model for free without their name on it for no reason.

show 1 reply
gadtflytoday at 3:26 AM

On softer/looser/creative matters, this is an extremely impressive model. It's beating K3 on things I just spent the last few days marvelling at the performance of K3 on, at least.

Visual reasoning is not great (unsurprising).

markasoftwaretoday at 4:30 AM

Anonymous unreleased models are made available on arena.ai all the time, it's not really news that one is on openrouter...

spdustintoday at 3:26 AM

Based on its indecisive and far-too-lengthy thinking traces when given complex instructions that span system and user messages, as well as a rudimentary stylometry (POS ratios in thinking traces, mainly) comparison with latest non-stealth models, this is almost certainly a GLM model.

show 1 reply
raincoletoday at 2:40 AM

Can someone enlighten me? I honestly don't get what it is or what it's for. Surely OpenRouter knows who the providers are?

show 2 replies
dozerlytoday at 1:39 AM

Yea, nice try there North Korea.

show 2 replies
raybbtoday at 1:47 AM

When a model is free like this what kind of rate limits are there?

show 1 reply
zb3today at 1:24 AM

We can know if this is Anthropic/OpenAI by testing the "guardrails" - absurd guardrails = it's them, reasonable/no guardrails = Chinese models..

(as a bonus - thinking forever = GLM)

show 1 reply
firlooptoday at 1:38 AM

I'm against stealth models—we should know what it is and see a model card with a list of safety considerations. Bit ridiculous of a practice to me.