GenAIExamples

Author	SHA1	Message	Date
ZePan110	785ffb9a1e	Update compose.yaml for ChatQnA (#1621 ) Update compose.yaml for ChatQnA Signed-off-by: ZePan110 <ze.pan@intel.com>	2025-03-07 09:19:39 +08:00
ZePan110	6ead1b12db	Enable ChatQnA model cache for docker compose test. (#1605 ) Enable ChatQnA model cache for docker compose test. Signed-off-by: ZePan110 <ze.pan@intel.com>	2025-03-05 11:30:04 +08:00
Artem Astafev	6abf7652e8	Fix ChatQnA ROCm compose Readme file and absolute path for ROCM CI test (#1159 ) Signed-off-by: Artem Astafev <a.astafev@datamonsters.com>	2025-02-27 15:26:45 +08:00
Wang, Kai Lawrence	284db982be	[ROCm] Fix the hf-token setting for TGI and TEI in ChatQnA (#1432 ) This PR is to correct the env variable names in chatqna example on ROCm platform passing to the docker container of TGI and TEI. For tgi, either HF_TOKEN and HUGGING_FACE_HUB_TOKEN could be parsed in TGI while HF_API_TOKEN can be parsed in TEI. TGI: https://github.com/huggingface/text-generation-inference/blob/main/router/src/server.rs#L1700C1-L1702C15 TEI: https://github.com/huggingface/text-embeddings-inference/blob/main/router/src/main.rs#L112 Signed-off-by: Wang, Kai Lawrence <kai.lawrence.wang@intel.com>	2025-01-21 14:22:39 +08:00
Wang, Kai Lawrence	3d3ac59bfb	[ChatQnA] Update the default LLM to llama3-8B on cpu/gpu/hpu (#1430 ) Update the default LLM to llama3-8B on cpu/nvgpu/amdgpu/gaudi for docker-compose deployment to avoid the potential model serving issue or the missing chat-template issue using neural-chat-7b. Slow serving issue of neural-chat-7b on ICX: #1420 Signed-off-by: Wang, Kai Lawrence <kai.lawrence.wang@intel.com>	2025-01-20 22:47:56 +08:00
Liang Lv	0f7e5a37ac	Adapt code for dataprep microservice refactor (#1408 ) https://github.com/opea-project/GenAIComps/pull/1153 Signed-off-by: lvliang-intel <liang1.lv@intel.com>	2025-01-20 20:37:03 +08:00
Letong Han	4cabd55778	Refactor Retrievers related Examples (#1387 ) Delete redundant retrievers docker image in docker_images_list.md. Refactor Retrievers related Examples READMEs. Change all of the comps/retrievers/xxx/xxx/Dockerfile path into comps/retrievers/src/Dockerfile. Fix the Examples CI issues of PR opea-project/GenAIComps#1138. Signed-off-by: letonghan <letong.han@intel.com>	2025-01-16 14:21:48 +08:00
Liang Lv	3ca78867eb	Update example code for embedding dependency moving to 3rd_party (#1368 ) Signed-off-by: lvliang-intel <liang1.lv@intel.com>	2025-01-10 15:36:58 +08:00
chen, suyue	5c7a5bd850	Update Code and README for GenAIComps Refactor (#1285 ) Signed-off-by: lvliang-intel <liang1.lv@intel.com> Signed-off-by: chensuyue <suyue.chen@intel.com> Signed-off-by: Xinyao Wang <xinyao.wang@intel.com> Signed-off-by: letonghan <letong.han@intel.com> Signed-off-by: ZePan110 <ze.pan@intel.com> Signed-off-by: WenjiaoYue <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>	2025-01-02 20:03:26 +08:00
Wang, Kai Lawrence	ac470421d0	Update the llm backend ports (#1172 ) Signed-off-by: Wang, Kai Lawrence <kai.lawrence.wang@intel.com>	2024-11-22 09:20:09 +08:00
Artem Astafev	6d3a017609	Add compose example for ChatQnA AMD ROCm deployment (#1122 ) Signed-off-by: Artem Astafev <a.astafev@datamonsters.com> Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>	2024-11-15 17:24:06 +08:00

11 Commits