🍌 Banana Serverless

This repo gives a basic framework for serving PygmalionAI's pygmalion-6b GPTJ model in production using Banana's serverless platform. Original model can be found here

Quickstart:

The repo is already set up to run a basic HuggingFace GPTJ model.

Run pip3 install -r requirements.txt to download dependencies.
Run python3 server.py to start the server.
Run python3 test.py in a different terminal session to test against it.

Make it your own:

Edit app.py to load and run your model.
Make sure to test with test.py!

if deploying using Docker:

Edit download.py (or the Dockerfile itself) with scripts download your custom model weights at build time.

Move to prod:

At this point, you have a functioning http server for your ML model. You can use it as is, or package it up with our provided Dockerfile and deploy it to your favorite container hosting provider!

If Banana is your favorite GPU hosting provider (and we sure hope it is), read on!

🍌

Deploy to Banana Serverless:

Log in to the Banana App
Select your customized repo for deploy!

It'll then be built from the dockerfile, optimized, then deployed on our Serverless GPU cluster and callable with any of our SDKs:

You can monitor buildtime and runtime logs by clicking the logs button in the model view on the Banana Dashboard](https://app.banana.dev)

Name		Name	Last commit message	Last commit date
Latest commit History 4 Commits
.gitignore		.gitignore
Dockerfile		Dockerfile
LICENSE		LICENSE
README.md		README.md
app.py		app.py
banana_config.json		banana_config.json
download.py		download.py
requirements.txt		requirements.txt
server.py		server.py
test.py		test.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

🍌 Banana Serverless

Quickstart:

Make it your own:

Move to prod:

🍌

Deploy to Banana Serverless:

Use Banana for scale.

About

Languages

License

lucataco/serverless-pygmalion-6b

Folders and files

Latest commit

History

Repository files navigation

🍌 Banana Serverless

Quickstart:

Make it your own:

Move to prod:

🍌

Deploy to Banana Serverless:

Use Banana for scale.

About

Topics

Resources

License

Stars

Watchers

Forks

Languages