Comments (5)
Issue-Label Bot is automatically applying the label feature_request
to this issue, with a confidence of 0.94. Please mark this comment with 👍 or 👎 to give our bot feedback!
Links: dashboard, app homepage and code for this bot.
from kserve.
/kind feature
from kserve.
Thanks for the feature! Given our data model (high level "Tensorflow" spec), do you think it would make most sense to swap different Tensorflow technologies with annotations?
Regarding:
"Maybe we also should investigate how to support co-serving multiple models in one serving CRD.".
I think this is an anti-pattern in KFServing. A KFServing Service is a "unit" of model serving. The idea of hosting multiple models in a single model server is extremely interesting and something I've been working on in my share time. I've been thinking we should control this via implementation, not interface, and allow users to specify an annotation like "enable-multitenancy".
I'll make a separate issue for this since I think it's a huge topic.
from kserve.
This is either implemented or open in separate issues
/close
from kserve.
@ellis-bigelow: Closing this issue.
In response to this:
This is either implemented or open in separate issues
/close
Instructions for interacting with me using PR comments are available here. If you have questions or suggestions related to my behavior, please file an issue against the kubernetes/test-infra repository.
from kserve.
Related Issues (20)
- Create github action to upload release yaml artifacts automatically
- Make MAX_GRPC_MESSAGE_LENGTH Configurable for Image Input Size Flexibility
- `NCCL` and `flash_atten` packages are missing in huggingface runtime
- ClusterRole permissions are too broadly scoped? HOT 1
- Add Oracle Cloud Infrastructure (OCI) Object Storage as a storage agent
- VirtualService regex match should be case insensitive
- Python SDK for KServe and Kubeflow Pipelines can not be installed at the same time HOT 2
- UnicodeDecodeError for grpcurl request with Bytes column in DataFrame HOT 7
- error setting up interface service HOT 2
- mlflow model cannot be loaded HOT 8
- stop using `gcr.io/kubebuilder/kube-rbac-proxy` before `18 March 2025` (image being deleted) HOT 1
- add Xinfernece ( an inference platform which integrated transformers, vllm, and llama.cpp as engines,) runtime for LLM Serving Runtime HOT 5
- Completion fails when echo is true with vLLM backend
- protobuf version conflict while trying to integrate with kfp HOT 2
- Client fails to list clusterservingruntimes HOT 2
- Not able to access torchserve custom metrics after deploying inference service on kserve
- The request to InferenceService is sent twice
- Getting timeout failed to failed to call webhook: Post "https://kserve-webhook-server-service.default.svc:443/mutate-serving-kserve-io-v1beta1-inferenceservice?timeout=10s" HOT 7
- Multi-Lora support
- fake client returns no kind "ClusterServingRuntimeList" is registered for version "serving/v1alpha1" HOT 1
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from kserve.