Featherless is a platform to use the very latest LLMs. We list thousands of models from Hugging Face and make them available serverless, helping you keep up with the hundreds of new models added daily without ever renting a GPU
No matter your application, find and use the state of the art LLMs with Featherless
Hello Product Hunters!
🪶 I'm super excited to launch Featherless.AI today!
A platform that allows quick access to all the top 🦙 models you see in Hugging Face 🤗 today. From the 8B to the merged 11B's and Qwen-2 72B's
I know it's daunting for folks to download and set up all these large models on GPUs to try them, one at a time. Also renting GPUs can be incredibly expensive as it's typically several dollars per hour. On the flip side, popular providers may not have the various finetune, or even weird model sizes that you would love to try.
That's why we built featherless AI. To eventually download and provide access to **all** of hugging face public models. Making all the various open-source AI models more accessible.
Starting with a simple $10 or $25 monthly plan, with unlimited usage (within the concurrency limit).
So anyone can use it personally, across any model, without worrying about token pricing in their day-to-day usage.
Also: as the team who also helps build open-source foundation models (hey RWKV folks!), we fully understand the concerns various folks have over data security and privacy. As such, featherless.ai has no logging of any of your message prompts and completion. Why? Because we are not interested in stealing your data to train our models.
So use it any way you want to, dun let someone else tell you how you should use your AI models.
🪄 The magic behind it?
At the heart of it, is a custom-built inference infrastructure, built by the team here from Recursal.AI : which was built to be able to dynamically hot-swap models in sub-seconds. Allowing us to rapidly autoscale and dynamically adjust our infrastructure based on what models are popular. Once we have the models downloaded into our cluster.
This allows us to provide more models, where previous providers have been limited in ensuring every model hosted has a dedicated GPU for it.
💬 In summary
🏃♂️ Run any of over 450+ huggingface models
🛠️ OpenAI compatible API, use your existing tools or client
📈 Unlimited usage (within concurrent usage)
🦙 Starting at $10/month for <15B models
🦅 To $25/month for 72B models
🎁 Special for PH: Signup with a subscription, and add referral `hello+producthunt@featherless.ai` for $10 off your next month bill
Feel free to ask me anything here on the product hunt launch!
And give it a try with a free trial, which allows you to chat with the models (up to a limited amount of messages) at
https://featherless.ai
@manisha_hr_ Thanks!
Its OpenAI API - so if your devs were already using OpenAI styled API for AI, they should be able to try this instantly (as its the same API)
Awesome. It's nice to be able to quickly try out and preview different llama models without having to deploy them on my own servers. Been meaning to the different variety of roleplay models for my dnd sessions. Are you planning to support other open-source models besides llama?
@taishiling - im glad you like being able to play with all the models.
Yes, we are currently downloading more models which would be coming online over the next few days.
We currently support RWKV and LLaMA based models. We plan to introduce Mistral MoE's next (to be confirmed), followed by potentially larger models.
The main reason we started with llama, is because it was the largest pool of all the popular models our initial users and community wanted to use.
But the goal remains: ALL huggingface models. One major group at a time
Report
Congrats on the launch! Featherless AI offers great access to Hugging Face models at a fair price Excited to see what it can do!
@nathan_wilce Thanks!
All the models are pre-downloaded to our clusters, and are on standby.
The GPUs spin's up and swap these models on the fly under seconds.
This was made possible with our custom built inference systems we built and optimized on.
Allowing us to keep cost low, and auto scale to actual usage.
PS: 450+ models takes up over 9TB of storage space, we are downloading 100+ more which will be coming online soon. (it takes time haha)
How does featherless tie into the novel models that you developed with recursal? Is it similar to OpenRouter? What are some of the top reasons to switch over to Featherless vs using other API's?
Report
Hunter
@the_esc
Expect to see the Eagle and RWKV models on featherless shortly - we targeted the Llama-3 based models first since it's more widely known and has a very rich and varied set of fine-tunes, but we definitely believe that the future belongs to RWKV :)
We're certainly similar to OpenRouter in providing an enormous range of models servelessly. We aren't listing other providers and we will likely complement OpenRouter by acting as a provider for sufficiently-popular-but-still-niche models.
The main reason to use featherless is to experiment and use a wider variety of fine-tunes than anywhere else.
@william_bowen4
Exactly! Useful when you want to test many models!
Report
Inoticed your pricing details did not mention anything about API access. Perhaps you could update it so non PH users can see it too :)
Congrats on the launch!
Featherless AI
Featherless AI
Featherless AI
Featherless AI
UI-licious
Featherless AI
Featherless AI
Featherless AI
Trellis
Featherless AI
Featherless AI
Featherless AI