What’s Changed: Hapuppy is a real, working API reseller, but it is not a safe home for money or chats you care about. Its own status page showed 6 of its 15 monitored models failing every check from October 1 to 3, 2026. Its privacy policy and terms also disagree about whether your messages get scanned.
Is Hapuppy legit? It is a real service that sells working API keys, and its staff have fixed broken payments by hand.
It is also a business with no company name, a roleplay Claude priced at 8% of Anthropic’s rate on its most popular pack, and a status page that contradicts its own homepage.
Hapuppy took off with SillyTavern roleplayers in mid-August 2026, after a well-known preset creator gave it a shoutout. Warning threads followed within a month.
One side says it works and costs almost nothing. The other side calls it a scam, and I don’t think either side has it right.
The useful answer sits in Hapuppy’s own documents. This guide covers its status data, its privacy terms, its pricing math and its refund rules, then gives you a ten-minute test for checking any reseller’s model.

Is Hapuppy Legit or Just Too Cheap to Last?
Hapuppy is legit in the narrow sense that it exists, takes payment and serves models, but it names no company, no address and no jurisdiction, and its own terms call it an “API relay service.”

What is an API relay: A middleman server that takes your API requests, forwards them to a model provider under its own accounts, and passes the replies back to you.
Hapuppy sells access to “106+ models” through one OpenAI-compatible endpoint, the kind of connection SillyTavern users paste into their API settings. Its terms of service say disputes fall under “the laws applicable in the jurisdiction where Hapuppy operates.” That jurisdiction is never named.
The domain is young. The registry record shows hapuppy.com was registered on February 10, 2026, through Alibaba Cloud’s HiChina registrar. Every account must also be tied to a Discord account before its key works, per Hapuppy’s verification guide.
None of that makes it a scam. When a buyer’s Google Pay charge went missing in August, Hapuppy staff credited the account by hand, and that buyer’s write-up concluded that “Hapuppy seems legit if somewhat overwhelmed.” What bothers me is that nobody is accountable for any of it, and the price makes no sense if the models were bought at list.
If the API side is the part you are tired of, there is a simpler route. People who want the roleplay without managing API keys and resellers tend to land on a hosted app like Nectar AI, where the characters and the model come in one account.
Why Does Hapuppy’s Status Page Show So Many Models Down?
Hapuppy’s status page shows models down because six of the 15 models it monitors failed all 100 checks between October 1 and October 3, 2026, while its homepage promises “99.9% uptime.”

Hapuppy runs a public status page for its models. An archived copy of its log records a check on every model each half hour, and the table counts the results from just after midnight UTC on October 1 to early on October 3:
| Model on Hapuppy | Checks passed (of 100) | Typical response time |
|---|---|---|
| Kimi K3 | 0 | No reply |
| GPT-6 Astra, GPT-5.6 Sol and Terra, DeepSeek V4 Flash and Pro | 0 each | No reply |
| Claude Opus 4.6 | 98 | About 2.2 seconds |
| Claude Sonnet 4.6 | 100 | About 1.6 seconds |
| GLM 5.3 | 97 | About 3.1 seconds |
| GLM 5.2, Kimi K2.6 | 100 each | Under 1 second |
Kimi K3 is the one that stings, because cheap Kimi K3 was the pitch that brought people in. The September Stay clear of Hapuppy thread called Kimi 3 “practically dead for days.”
A check that fails 100 times in a row can also mean a model name Hapuppy retired without updating its monitor. The dated DeepSeek versions passed nearly every check while the undated names failed. I treat a status page as the operator’s own admission, though, and this one admits a big slice of the catalog was not answering.
The uptime promise also gets no backing from the terms. Section 8 says “We do not offer SLA guarantees or service credits for downtime,” and section 9 lets Hapuppy remove models “without prior notice.”
The outages are not new, either. In May 2026, a developer filed a bug report showing Hapuppy returning HTTP 500 errors on roughly 30 to 40% of Gemini 3.5 Flash requests. Its final streaming chunk was also malformed, so the developer’s app read every reply as an error.
Does Hapuppy Read Your Chats?
Hapuppy’s own documents give two answers: its privacy policy says it does not store, log or inspect your requests, while its terms say submitted content “may be scanned, filtered, or temporarily retained.”
Both documents carry the same date, March 11, 2026. Here is what each one says:
| Hapuppy document | What it says |
|---|---|
| Privacy policy | “We do not store, log, or inspect the content of your API requests or responses.” |
| Terms of service | “Submitted content may be scanned, filtered, or temporarily retained for review and legal compliance.” |
| Privacy policy | “We do not use third-party tracking cookies, advertising cookies, or analytics cookies.” |
| Homepage | Its page code preloads Google’s analytics script, tag G-XNTTCN4L4L |
The terms even send readers to the privacy policy “for the full description of our moderation practices.” The privacy policy has no moderation section at all.
The scanning half has backing beyond the terms. Commenters in the same warning thread describe an announcement that every output runs through Hapuppy’s own filtering model, which Hapuppy says is there to catch content involving minors.
A filter for illegal content is a fair thing to run. My problem is a privacy policy that says the opposite, because a service that will not describe what it keeps should not hold a story you would hate to see read aloud.
The paperwork has smaller cracks too. The privacy policy tells readers they can “cancel your subscription at any time through Creem,” Hapuppy’s payment processor.
The homepage advertises billing “from $9.99/mo.” Hapuppy’s FAQ says the packs are one-time purchases with “nothing to cancel.”
How Can Hapuppy Sell Claude So Cheap?
Hapuppy’s site never explains its Claude prices, and they do not add up: its roleplay Opus costs $2.00 per million output tokens against Anthropic’s $25.00, and caching cannot discount output.
The math comes straight from Hapuppy’s own pricing card, which sells 120 million credits for $19.99 and counts 1 million credits as $5.00 of Claude Opus 4.7 input. That $5 matches Anthropic’s published rate per million input tokens.
That works out to about $0.17 per million credits. Here is what Claude Opus 4.7 costs per million tokens in each case:
| Claude Opus 4.7 through | Input per 1M tokens | Output per 1M tokens | Share of Anthropic’s price |
|---|---|---|---|
| Anthropic directly | $5.00 | $25.00 | 100% |
| Hapuppy roleplay version, $9.99 pack | $0.80 | $4.00 | 16% |
| Hapuppy roleplay version, $19.99 pack | $0.40 | $2.00 | 8% |
| Hapuppy no-roleplay version, $19.99 pack | $0.17 | $0.83 | About 3% |
Roleplayers land in the expensive rows. Hapuppy’s Claude docs say its free Claude models with plain names, such as claude-opus-4-7, “do NOT support role-play.” The paid versions named with a hapuppy/ prefix are described as “the same underlying models, but configured to allow role-play (RP) scenarios that the standard variants refuse.”
So the “97%” on Hapuppy’s pricing card is a price most SillyTavern users never pay. Its credits guide prices the roleplay version at 2.4 times the no-roleplay rate on both input and output, and a reconfigured Claude is the part of this setup I trust least.
Hapuppy’s site and docs never say where the discount comes from. The preset creator who promoted it said in the payment thread that the low price comes from “zero overhead profit,” with the service “ran like a hobby rather than a business.”
ChinaTalk’s report on API relays describes the usual sources instead. They include farmed free credits, resold quota, one $200 Claude Max plan shared among many users, and accounts bought with stolen cards.
That report puts typical relay prices at 70 to 90% below official rates. It also lists the logs as a product, since every prompt and reply passes through the operator’s servers.
Hapuppy’s site offers a Chinese-language version and shows its $19.99 pack as ¥139. None of that proves how it sources tokens, and I would not claim it does. A hobby still pays its upstream bill, though, and at the published price every roleplay reply would cost Hapuppy far more than it charges.
Cheap providers change their deals fast even when they are honest. Janitor AI users learned that when Chutes raised its prices in June 2026.
Is Hapuppy’s Claude Really Claude?
Nobody outside Hapuppy can prove which model answers your request, and that is the real risk: a CISPA Helmholtz Center audit of unofficial API resellers found model identity checks failing in 45.83% of fingerprint tests.
A commenter in the Stay clear thread wrote that “Hapuppy’s claude is fake also” and “just GLM or some other model.” GLM is a much cheaper model family from China’s Zhipu AI. The comment comes with no test behind it.
Model swapping is a documented practice, though. The CISPA paper “Real Money, Fake Models” found 17 shadow APIs used in 187 academic papers, and its audit measured performance gaps of up to 47.21% against the official APIs.
Roleplayers watched it happen in September. CrofAI, another cheap provider in the same circles, was exposed as an OpenRouter wrapper routing requests to cheaper models at up to 20 times markup, and it wiped its online presence within hours.
I won’t call Hapuppy’s Claude fake on one comment’s word. You can run your own check in about ten minutes.
How Do You Test Whether a Reseller Serves the Real Model?
The cleanest test is a token count comparison: send the identical prompt to a trusted source for the model and to the reseller, then compare the input token totals each one reports.
Different model families split text into tokens differently. Anthropic’s pricing page also says Claude 4.7 and later use a tokenizer that makes about 30% more tokens from the same text, so even a mislabeled Claude version shows up.
Here is the order I run it in:
- Copy about 1,000 words of your own story into a test prompt.
- Send it once through a source you trust, such as Anthropic’s API or OpenRouter with the Anthropic provider pinned.
- Send the identical prompt through the reseller, with the same settings and temperature 0.
- Compare the input token totals in each raw response.
- Repeat with a 500-word version and see how the gap changes.
Test prompt: "Continue this scene in exactly 150 words: [paste your story]"
Trusted route input tokens: A
Reseller route input tokens: B
Same gap at 500 and 1,000 words means hidden instructions were added
Gap grows with the text means a different tokenizer, so a different model familyOn Hapuppy’s roleplay versions, a fixed gap is expected, because its docs say those models are “configured” for roleplay. A reseller can also rewrite these numbers, so a match proves little. A mismatch that grows with your text is much harder to explain away.
Can You Get a Refund From Hapuppy?
Hapuppy refunds only a first purchase, within 7 days and under 300,000 credits of use, which is about 12 replies in a typical SillyTavern chat on its roleplay Opus.
Example scenario: A SillyTavern chat with 8,000 tokens of context and a 400-token reply costs 8,000 × 2.4 plus 400 × 12 credits on hapuppy/claude-opus-4-7. That is 24,000 credits a reply. A dozen replies reaches 288,000 credits, and the thirteenth takes you past the refund limit.
Hapuppy’s terms add two more conditions. Refunds apply to first-time packs only, and Creem’s processing fee of 3.9% plus $0.40 comes off the top.
Hapuppy’s credits guide states the rule more loosely, as not having “consumed a significant portion” of the pack. The terms are the stricter version, so plan around 300,000.
Payment is the other weak spot. In the August payment thread, Creem’s support bot told Hapuppy that a buyer’s Google Pay charge from 2026/08/14 was fraud because “the year 2026 is in the future.” Hapuppy’s support credited the buyer anyway.
The same buyer posted days later that “All my remaining 115 million credits are gone.” That turned out to be a dashboard display bug and the credits came back, but the buyer still would not recommend the service unless cheap tokens are your only priority.
When I top up a reseller like this, it is the smallest pack on a card I can cancel, never the debit card tied to my main account. That rule costs nothing on the days the service works.
Should You Use Hapuppy With SillyTavern?
Use Hapuppy with SillyTavern only as a throwaway experiment: a small top-up, no private material, and a second provider ready, because its models and policies can change without notice.
| Your situation | Better pick | Why |
|---|---|---|
| You want the cheapest tokens and can lose $10 | Hapuppy’s $9.99 pack | Very cheap Claude, but no uptime guarantee and a refund window of about a dozen replies |
| You want to choose which provider serves each reply | OpenRouter | Lets you pin a named provider for a model, so you know whose servers answered |
| You want to spend nothing while you learn | A free model setup | The free OpenRouter setup costs nothing to try |
| You want the story, not the API | A hosted companion app | Nothing to configure and no Discord binding |
OpenRouter is the transparent option, and it is not a small one. It claims 8 million users and 100 trillion tokens a month, according to TechCrunch’s report on its $1.3 billion valuation.
The full provider comparison, including NanoGPT and when a cheap reseller makes sense, is in the SillyTavern models guide. If you still want to try Hapuppy, keep the risk small:
- Buy the $9.99 pack, not the $36.99 one, until it has worked for you for a week.
- Pay with a card you can cancel, never your main debit card.
- Check the status page before you build a preset around any model.
- Save a second provider in SillyTavern so a dead model does not end your evening.
- Keep anything you would not want read off it.
If the API side is the part wearing you out, Candy AI is the hosted option I point people to when they are done managing keys. Its free tier is only 5 messages, so treat it as a test drive rather than a plan.
Quick Takeaways
- Hapuppy is a working service run by an unnamed operator, and its staff have fixed payments by hand.
- Its status page showed 6 of 15 models failing every check from October 1 to 3, 2026, against a 99.9% uptime claim.
- Its privacy policy says it never inspects requests, while its terms say content may be scanned and retained.
- Roleplay Opus costs $2.00 per million output tokens on Hapuppy against Anthropic’s $25, a gap Hapuppy never explains.
- If you try it, spend $9.99, keep private stories off it, and run the token count test first.

The status table is the most telling part: six models failing all 100 checks while the homepage still advertises 99.9% uptime suggests the uptime claim is marketing rather than a measured figure. I’d be curious whether those failures cluster around specific upstream providers, since that would hint at routing or account issues rather than random outages. Either way, a public log that contradicts the front page is a useful signal for anyone weighing prepaid credits.