Comments (6)
How do you feel about annotations?
from kserve.
@ellis-bigelow any further thoughts on this? By checking this page, looks a bunch of envs are needed
https://github.com/kubeflow/examples/tree/master/mnist#using-s3, not sure how to use annotation.
from kserve.
Here's the flow I had in mind:
- User must have pre-created secret
- User provides an annotation referencing the secret:
"serving.kubeflow.org/kfservice-s3-model-secret: secretRef"
- In our controller, if the user provides that annotation, we mount those to the kfconfigurations
I only worry a little bit that our default + canary model is a little weird. What happens if a user wants to canary a new secret -- they'll take an outage. Perhaps we can qualify the annotations?
We may want to not specify s3 and just mount the secret to all relevant locations for gcs/s3/blob etc
Alternatively, we could use a k8s service account and mount them in the right place in the pod for s3/gcs/blob. What do you think?
from kserve.
With our model once secretRef annotation is changed both default and canary get updated, so user will take an outage for a bad secret rollout. Maybe we can support envs on default/canary spec and pull secrets as envs which works for s3 but not sure for other storage.
We may want to not specify s3 and just mount the secret to all relevant locations for gcs/s3/blob etc
Alternatively, we could use a k8s service account and mount them in the right place in the pod for s3/gcs/blob. What do you think?
Can you elaborate these more not sure if I understand this.
from kserve.
/assign @yuzisun
from kserve.
Discussed with @ellis-bigelow as a first pass we can add service account to the spec to which user adds the secrets for gcs/s3 or other identities.
from kserve.
Related Issues (20)
- Make MAX_GRPC_MESSAGE_LENGTH Configurable for Image Input Size Flexibility
- `NCCL` and `flash_atten` packages are missing in huggingface runtime
- ClusterRole permissions are too broadly scoped? HOT 1
- Add Oracle Cloud Infrastructure (OCI) Object Storage as a storage agent
- VirtualService regex match should be case insensitive
- Python SDK for KServe and Kubeflow Pipelines can not be installed at the same time HOT 2
- UnicodeDecodeError for grpcurl request with Bytes column in DataFrame HOT 7
- error setting up interface service HOT 2
- mlflow model cannot be loaded HOT 8
- stop using `gcr.io/kubebuilder/kube-rbac-proxy` before `18 March 2025` (image being deleted) HOT 1
- add Xinfernece ( an inference platform which integrated transformers, vllm, and llama.cpp as engines,) runtime for LLM Serving Runtime HOT 5
- Completion fails when echo is true with vLLM backend
- protobuf version conflict while trying to integrate with kfp HOT 2
- Client fails to list clusterservingruntimes HOT 2
- Not able to access torchserve custom metrics after deploying inference service on kserve
- The request to InferenceService is sent twice
- Getting timeout failed to failed to call webhook: Post "https://kserve-webhook-server-service.default.svc:443/mutate-serving-kserve-io-v1beta1-inferenceservice?timeout=10s" HOT 7
- Multi-Lora support
- fake client returns no kind "ClusterServingRuntimeList" is registered for version "serving/v1alpha1" HOT 1
- duplicated hosts error after configuring the additional domains HOT 1
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from kserve.