René's URL Explorer Experiment


Title: GitHub - intel/ipex-llm: Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discrete GPU such as Arc, Flex and Max); seamlessly integrate with llama.cpp, Ollama, HuggingFace, LangChain, LlamaIndex, vLLM, DeepSpeed, Axolotl, etc. · GitHub

Open Graph Title: GitHub - intel/ipex-llm: Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discrete GPU such as Arc, Flex and Max); seamlessly integrate with llama.cpp, Ollama, HuggingFace, LangChain, LlamaIndex, vLLM, DeepSpeed, Axolotl, etc.

X Title: GitHub - intel/ipex-llm: Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discrete GPU such as Arc, Flex and Max); seamlessly integrate with llama.cpp, Ollama, HuggingFace, LangChain, LlamaIndex, vLLM, DeepSpeed, Axolotl, etc.

Description: Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discrete GPU such as Arc, Flex and Max); seamlessly integrate with llama.cpp, Ollama, HuggingFace, LangChain, LlamaIndex, vLLM, DeepSpeed, Axolotl, etc. - intel/ipex-llm

Open Graph Description: Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discr...

X Description: Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discr...

Opengraph URL: https://github.com/intel/ipex-llm

X: @github

direct link

Domain: github.com

route-pattern/:user_id/:repository
route-controllerfiles
route-actiondisambiguate
fetch-noncev2:a1ef4059-01fd-c470-0075-4bf2f9bdc624
current-catalog-service-hashf3abb0cc802f3d7b95fc8762b94bdcb13bf39634c40c357301c4aa1d67a256fb
request-id8CAA:241082:22E3A2E:2EDBCC9:6A653B70
html-safe-nonce9ce4c381f94bf23168085d335b7edbac8cf227627a1170573cfe8788c5ae2820
visitor-payloadeyJyZWZlcnJlciI6IiIsInJlcXVlc3RfaWQiOiI4Q0FBOjI0MTA4MjoyMkUzQTJFOjJFREJDQzk6NkE2NTNCNzAiLCJ2aXNpdG9yX2lkIjoiMTQ2MjA5ODM2NTAzMTcyNTkzNiIsInJlZ2lvbl9lZGdlIjoiaWFkIiwicmVnaW9uX3JlbmRlciI6ImlhZCJ9
visitor-hmac54be23a22129f97207741c0fe4f93761a5c0eafe13868284e6e5b9c7025000e4
hovercard-subject-tagrepository:66823715
github-keyboard-shortcutsrepository,copilot
google-site-verificationApib7-x98H0j5cPqHWwSMm6dNU4GmODRoqxLiDzdx9I
octolytics-urlhttps://collector.github.com/github/collect
analytics-location//
fb:app_id1401488693436528
apple-itunes-appapp-id=1477376905, app-argument=https://github.com/intel/ipex-llm
twitter:imagehttps://opengraph.githubassets.com/22558831803b2778e24b04c97d1af1d56689f92eff957a78f5a933a033345e5a/intel/ipex-llm
twitter:cardsummary_large_image
og:imagehttps://opengraph.githubassets.com/22558831803b2778e24b04c97d1af1d56689f92eff957a78f5a933a033345e5a/intel/ipex-llm
og:image:altAccelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discr...
og:image:width1200
og:image:height600
og:site_nameGitHub
og:typeobject
hostnamegithub.com
expected-hostnamegithub.com
None52c76df668885aaff23b50bdca1fa1ea44ac9c1553e888ebc70ff1e4daa4625b
turbo-cache-controlno-cache
go-importgithub.com/intel/ipex-llm git https://github.com/intel/ipex-llm.git
octolytics-dimension-user_id17888862
octolytics-dimension-user_loginintel
octolytics-dimension-repository_id66823715
octolytics-dimension-repository_nwointel/ipex-llm
octolytics-dimension-repository_publictrue
octolytics-dimension-repository_is_forkfalse
octolytics-dimension-repository_network_root_id66823715
octolytics-dimension-repository_network_root_nwointel/ipex-llm
turbo-body-classeslogged-out env-production page-responsive
disable-turbofalse
browser-stats-urlhttps://api.github.com/_private/browser/stats
browser-errors-urlhttps://api.github.com/_private/browser/errors
release309153364422b3c499922d1a2a6404910a58ed8e
ui-targetfull
theme-color#1e2327
color-schemelight dark

Links:

Skip to contenthttps://github.com/intel/ipex-llm#start-of-content
https://github.com/
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2Fintel%2Fipex-llm
GitHub CopilotWrite better code with AIhttps://github.com/features/copilot
GitHub Copilot appDirect agents from issue to mergehttps://github.com/features/ai/github-app
MCP RegistryNewIntegrate external toolshttps://github.com/mcp
ActionsAutomate any workflowhttps://github.com/features/actions
CodespacesInstant dev environmentshttps://github.com/features/codespaces
IssuesPlan and track workhttps://github.com/features/issues
Code ReviewManage code changeshttps://github.com/features/code-review
Code QualityEnforce quality at mergehttps://github.com/features/code-quality
GitHub Advanced SecurityFind and fix vulnerabilitieshttps://github.com/security/advanced-security
Code securitySecure your code as you buildhttps://github.com/security/advanced-security/code-security
Secret protectionStop leaks before they starthttps://github.com/security/advanced-security/secret-protection
Why GitHubhttps://github.com/why-github
Documentationhttps://docs.github.com
Bloghttps://github.blog
Changeloghttps://github.blog/changelog
Marketplacehttps://github.com/marketplace
View all featureshttps://github.com/features
Enterpriseshttps://github.com/enterprise
Small and medium teamshttps://github.com/team
Startupshttps://github.com/enterprise/startups
Nonprofitshttps://github.com/solutions/industry/nonprofits
App Modernizationhttps://github.com/solutions/use-case/app-modernization
DevSecOpshttps://github.com/solutions/use-case/devsecops
DevOpshttps://github.com/solutions/use-case/devops
CI/CDhttps://github.com/solutions/use-case/ci-cd
View all use caseshttps://github.com/solutions/use-case
Healthcarehttps://github.com/solutions/industry/healthcare
Financial serviceshttps://github.com/solutions/industry/financial-services
Manufacturinghttps://github.com/solutions/industry/manufacturing
Governmenthttps://github.com/solutions/industry/government
View all industrieshttps://github.com/solutions/industry
View all solutionshttps://github.com/solutions
AIhttps://github.com/resources/articles?topic=ai
Software Developmenthttps://github.com/resources/articles?topic=software-development
DevOpshttps://github.com/resources/articles?topic=devops
Securityhttps://github.com/resources/articles?topic=security
View all topicshttps://github.com/resources/articles
Customer storieshttps://github.com/customer-stories
Events & webinarshttps://github.com/resources/events
Ebooks & reportshttps://github.com/resources/whitepapers
Business insightshttps://github.com/solutions/executive-insights
GitHub Skillshttps://skills.github.com
Documentationhttps://docs.github.com
Customer supporthttps://support.github.com
Community forumhttps://github.com/orgs/community/discussions
Trust centerhttps://github.com/trust-center
Partnershttps://github.com/partners
View all resourceshttps://github.com/resources
GitHub SponsorsFund open source developershttps://github.com/open-source/sponsors
Security Labhttps://securitylab.github.com
Maintainer Communityhttps://maintainers.github.com
Acceleratorhttps://github.com/open-source/accelerator
GitHub Starshttps://stars.github.com
Archive Programhttps://archiveprogram.github.com
Topicshttps://github.com/topics
Trendinghttps://github.com/trending
Collectionshttps://github.com/collections
Enterprise platformAI-powered developer platformhttps://github.com/enterprise
GitHub Advanced SecurityEnterprise-grade security featureshttps://github.com/security/advanced-security
Copilot for BusinessEnterprise-grade AI featureshttps://github.com/features/copilot/copilot-business
Premium SupportEnterprise-grade 24/7 supporthttps://github.com/enterprise/premium-support
Pricinghttps://github.com/pricing
Search syntax tipshttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
documentationhttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2Fintel%2Fipex-llm
Sign up https://github.com/signup?ref_cta=Sign+up&ref_loc=header+logged+out&ref_page=%2F%3Cuser-name%3E%2F%3Crepo-name%3E&source=header-repo&source_repo=intel%2Fipex-llm
Reloadhttps://github.com/intel/ipex-llm
Reloadhttps://github.com/intel/ipex-llm
Reloadhttps://github.com/intel/ipex-llm
Please reload this pagehttps://github.com/intel/ipex-llm
intel https://github.com/intel
ipex-llmhttps://github.com/intel/ipex-llm
Notifications https://github.com/login?return_to=%2Fintel%2Fipex-llm
Fork 1.4k https://github.com/login?return_to=%2Fintel%2Fipex-llm
Star 8.9k https://github.com/login?return_to=%2Fintel%2Fipex-llm
Code https://github.com/intel/ipex-llm
Issues 1.2k https://github.com/intel/ipex-llm/issues
Pull requests 271 https://github.com/intel/ipex-llm/pulls
Discussions https://github.com/intel/ipex-llm/discussions
Actions https://github.com/intel/ipex-llm/actions
Projects https://github.com/intel/ipex-llm/projects
Wiki https://github.com/intel/ipex-llm/wiki
Security and quality 0 https://github.com/intel/ipex-llm/security
Insights https://github.com/intel/ipex-llm/pulse
Code https://github.com/intel/ipex-llm
Issues https://github.com/intel/ipex-llm/issues
Pull requests https://github.com/intel/ipex-llm/pulls
Discussions https://github.com/intel/ipex-llm/discussions
Actions https://github.com/intel/ipex-llm/actions
Projects https://github.com/intel/ipex-llm/projects
Wiki https://github.com/intel/ipex-llm/wiki
Security and quality https://github.com/intel/ipex-llm/security
Insights https://github.com/intel/ipex-llm/pulse
https://github.com/intel/ipex-llm
Brancheshttps://github.com/intel/ipex-llm/branches
Tagshttps://github.com/intel/ipex-llm/tags
https://github.com/intel/ipex-llm/branches
https://github.com/intel/ipex-llm/tags
4,113 Commitshttps://github.com/intel/ipex-llm/commits/main/
https://github.com/intel/ipex-llm/commits/main/
.githubhttps://github.com/intel/ipex-llm/tree/main/.github
.githubhttps://github.com/intel/ipex-llm/tree/main/.github
appshttps://github.com/intel/ipex-llm/tree/main/apps
appshttps://github.com/intel/ipex-llm/tree/main/apps
docker/llmhttps://github.com/intel/ipex-llm/tree/main/docker/llm
docker/llmhttps://github.com/intel/ipex-llm/tree/main/docker/llm
docs/mddocshttps://github.com/intel/ipex-llm/tree/main/docs/mddocs
docs/mddocshttps://github.com/intel/ipex-llm/tree/main/docs/mddocs
python/llmhttps://github.com/intel/ipex-llm/tree/main/python/llm
python/llmhttps://github.com/intel/ipex-llm/tree/main/python/llm
.gitignorehttps://github.com/intel/ipex-llm/blob/main/.gitignore
.gitignorehttps://github.com/intel/ipex-llm/blob/main/.gitignore
.readthedocs.ymlhttps://github.com/intel/ipex-llm/blob/main/.readthedocs.yml
.readthedocs.ymlhttps://github.com/intel/ipex-llm/blob/main/.readthedocs.yml
LICENSEhttps://github.com/intel/ipex-llm/blob/main/LICENSE
LICENSEhttps://github.com/intel/ipex-llm/blob/main/LICENSE
README.mdhttps://github.com/intel/ipex-llm/blob/main/README.md
README.mdhttps://github.com/intel/ipex-llm/blob/main/README.md
README.zh-CN.mdhttps://github.com/intel/ipex-llm/blob/main/README.zh-CN.md
README.zh-CN.mdhttps://github.com/intel/ipex-llm/blob/main/README.zh-CN.md
SECURITY.mdhttps://github.com/intel/ipex-llm/blob/main/SECURITY.md
SECURITY.mdhttps://github.com/intel/ipex-llm/blob/main/SECURITY.md
pyproject.tomlhttps://github.com/intel/ipex-llm/blob/main/pyproject.toml
pyproject.tomlhttps://github.com/intel/ipex-llm/blob/main/pyproject.toml
READMEhttps://github.com/intel/ipex-llm
Code of conducthttps://github.com/intel/ipex-llm
Contributinghttps://github.com/intel/ipex-llm
Apache-2.0 licensehttps://github.com/intel/ipex-llm
Securityhttps://github.com/intel/ipex-llm
https://github.com/intel/ipex-llm#this-project-is-archived
https://github.com/intel/ipex-llm#-intel-llm-library-for-pytorch
中文https://github.com/intel/ipex-llm/blob/main/README.zh-CN.md
GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_windows_gpu.md
NPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/npu_quickstart.md
1https://github.com/intel/ipex-llm#user-content-fn-1-e9586e16fae8e38ed833d96c96f33351
llama.cpphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md
Ollamahttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_portable_zip_quickstart.md
vLLMhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/vLLM_quickstart.md
HuggingFace transformershttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace
LangChainhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LangChain
LlamaIndexhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LlamaIndex
Text-Generation-WebUIhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/webui_quickstart.md
DeepSpeed-AutoTPhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Deepspeed-AutoTP
FastChathttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/fastchat_quickstart.md
Axolotlhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/axolotl_quickstart.md
HuggingFace PEFThttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning
HuggingFace TRLhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/DPO
AutoGenhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Applications/autogen
ModeScopehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/ModelScope-Models
herehttps://github.com/intel/ipex-llm#verified-models
https://github.com/intel/ipex-llm#latest-update-
FlashMoEhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/flashmoe_quickstart.md
Ollama Portable Ziphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_portable_zip_quickstart.md
llama.cpp Portable Ziphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md
PyTorch 2.6https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_pytorch26_gpu.md
llama.cpp Portable Ziphttps://github.com/intel/ipex-llm/issues/12963#issuecomment-2724032898
llama.cpp Portable Ziphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md#flashmoe-for-deepseek-v3r1
llama.cpp Portable Ziphttps://github.com/ipex-llm/ipex-llm/releases/tag/v2.3.0-nightly
Windowshttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md#windows-quickstart
Linuxhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md#linux-quickstart
Windowshttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llama_cpp_npu_portable_zip_quickstart.md
Ollama Portable Ziphttps://github.com/ipex-llm/ipex-llm/releases/tag/v2.3.0-nightly
Windowshttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_portable_zip_quickstart.md#windows-quickstart
Linuxhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_portable_zip_quickstart.md#linux-quickstart
vLLM 0.6.6https://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/vllm_docker_quickstart.md
B580https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/bmg_quickstart.md
Ollama 0.5.4https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_quickstart.md
NPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/npu_quickstart.md
vLLM 0.6.2https://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/vllm_docker_quickstart.md
herehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/graphrag_quickstart.md
StableDiffusionhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/StableDiffusion
Phi-3-Visionhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/phi-3-vision
Qwen-VLhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/qwen-vl
morehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/More-Data-Types
herehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels
inferencehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Pipeline-Parallel-Inference
GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ragflow_quickstart.md
herehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/axolotl_quickstart.md
imageshttps://github.com/intel/ipex-llm#docker
one commandhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_windows_gpu.md#install-ipex-llm
herehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/open_webui_with_ollama_quickstart.md
herehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llama3_llamacpp_ollama_quickstart.md
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/llama3
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/llama3
llama.cpphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llama_cpp_quickstart.md
ollamahttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_quickstart.md
herehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/bigdl_llm_migration.md
herehttps://github.com/intel-analytics/bigdl-2.x
ModelScopehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/ModelScope-Models
魔搭https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/ModelScope-Models
IQ2https://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GGUF-IQ2
Text-Generation-WebUIhttps://github.com/intel-analytics/text-generation-webui
Self-Speculative Decodinghttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Inference/Self_Speculative_Decoding.md
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Speculative-Decoding
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Speculative-Decoding
LoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/LoRA
QLoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QLoRA
DPOhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/DPO
QA-LoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QA-LoRA
ReLoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/ReLora
QLoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QLoRA
Standford-Alpacahttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QLoRA/alpaca-qlora
herehttps://www.intel.com/content/www/us/en/developer/articles/technical/finetuning-llms-on-intel-gpus-using-bigdl-llm.html
ReLoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/ReLora
"ReLoRA: High-Rank Training Through Low-Rank Updates"https://arxiv.org/abs/2307.05695
Mixtral-8x7Bhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/mixtral
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/mixtral
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/mixtral
QA-LoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QA-LoRA
"QA-LoRA: Quantization-Aware Low-Rank Adaptation of Large Language Models"https://arxiv.org/abs/2309.14717
FP8 and FP4 inferencehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/More-Data-Types
GGUFhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GGUF
AWQhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/AWQ
GPTQhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GPTQ
vLLM continuous batchinghttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/vLLM-Serving
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/vLLM-Serving
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/vLLM-Serving
QLoRA finetuninghttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QLoRA
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QLoRA
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/QLoRA-FineTuning
FastChat servinghttps://github.com/intel/ipex-llm/blob/main/python/llm/src/ipex_llm/llm/serving
Intel GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU
tutorialhttps://github.com/intel-analytics/ipex-llm-tutorial
https://github.com/intel/ipex-llm#ipex-llm-demo
https://llm-assets.readthedocs.io/en/latest/_images/mtl_mistral-7B_q4_k_m_ollama.gif
https://llm-assets.readthedocs.io/en/latest/_images/npu_llama3.2-3B.gif
https://llm-assets.readthedocs.io/en/latest/_images/2arc_DeepSeek-R1-Distill-Qwen-32B-Q4_K_M.gif
https://llm-assets.readthedocs.io/en/latest/_images/FlashMoE-Qwen3-235B.gif
Ollama (Mistral-7B, Q4_K) https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_portable_zip_quickstart.md
HuggingFace (Llama3.2-3B, SYM_INT4)https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/npu_quickstart.md
llama.cpp (DeepSeek-R1-Distill-Qwen-32B, Q4_K)https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md
FlashMoE (Qwen3MoE-235B, Q4_K) https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/flashmoe_quickstart.md
https://github.com/intel/ipex-llm#ipex-llm-performance
1https://github.com/intel/ipex-llm#user-content-fn-1-e9586e16fae8e38ed833d96c96f33351
[2]https://www.intel.com/content/www/us/en/developer/articles/technical/accelerate-meta-llama3-with-intel-ai-solutions.html
[3]https://www.intel.com/content/www/us/en/developer/articles/technical/accelerate-microsoft-phi-3-models-intel-ai-soln.html
[4]https://www.intel.com/content/www/us/en/developer/articles/technical/intel-ai-solutions-accelerate-alibaba-qwen2-llms.html
https://llm-assets.readthedocs.io/en/latest/_images/MTL_perf.jpg
https://llm-assets.readthedocs.io/en/latest/_images/Arc_perf.jpg
Benchmarking Guidehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/benchmark_quickstart.md
https://github.com/intel/ipex-llm#model-accuracy
herehttps://github.com/intel-analytics/ipex-llm/tree/main/python/llm/dev/benchmark/perplexity
https://github.com/intel/ipex-llm#ipex-llm-quickstart
https://github.com/intel/ipex-llm#use
Ollamahttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ollama_portable_zip_quickstart.md
llama.cpphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llamacpp_portable_zip_gpu_quickstart.md
Arc B580https://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/bmg_quickstart.md
NPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/npu_quickstart.md
llama.cpphttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/llama_cpp_npu_portable_zip_quickstart.md
PyTorch/HuggingFacehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_windows_gpu.md
Windowshttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_windows_gpu.md
Linuxhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_linux_gpu.md
vLLMhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/vLLM_quickstart.md
GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/vllm_docker_quickstart.md
CPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/vllm_cpu_docker_quickstart.md
FastChathttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/fastchat_quickstart.md
Serving on multiple Intel GPUshttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/deepspeed_autotp_fastapi_quickstart.md
Text-Generation-WebUIhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/webui_quickstart.md
Axolotlhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/axolotl_quickstart.md
Benchmarkinghttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/benchmark_quickstart.md
https://github.com/intel/ipex-llm#docker
GPU Inference in C++https://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/docker_cpp_xpu_quickstart.md
GPU Inference in Pythonhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/docker_pytorch_inference_gpu.md
vLLM on GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/vllm_docker_quickstart.md
vLLM on CPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/vllm_cpu_docker_quickstart.md
FastChat on GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/fastchat_docker_quickstart.md
VSCode on GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/DockerGuides/docker_run_pytorch_inference_in_vscode.md
https://github.com/intel/ipex-llm#applications
GraphRAGhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/graphrag_quickstart.md
RAGFlowhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/ragflow_quickstart.md
LangChain-Chatchathttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/chatchat_quickstart.md
Coding copilothttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/continue_quickstart.md
Open WebUIhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/open_webui_with_ollama_quickstart.md
PrivateGPThttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/privateGPT_quickstart.md
Dify platformhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/dify_quickstart.md
https://github.com/intel/ipex-llm#install
Windows GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_windows_gpu.md
Linux GPUhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/install_linux_gpu.md
full installation guidehttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Overview/install.md
https://github.com/intel/ipex-llm#code-examples
https://github.com/intel/ipex-llm#low-bit-inference
INT4 inferencehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model
FP8/FP6/FP4 inferencehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/More-Data-Types
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/More-Data-Types
INT8 inferencehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/More-Data-Types
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/More-Data-Types
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/More-Data-Types
INT2 inferencehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GGUF-IQ2
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GGUF-IQ2
https://github.com/intel/ipex-llm#fp16bf16-inference
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Speculative-Decoding
self-speculative decodinghttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Inference/Self_Speculative_Decoding.md
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Speculative-Decoding
self-speculative decodinghttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Inference/Self_Speculative_Decoding.md
https://github.com/intel/ipex-llm#distributed-inference
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Pipeline-Parallel-Inference
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Deepspeed-AutoTP
https://github.com/intel/ipex-llm#save-and-load
Low-bit modelshttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Save-Load
GGUFhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GGUF
AWQhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/AWQ
GPTQhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Advanced-Quantizations/GPTQ
https://github.com/intel/ipex-llm#finetuning
GPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning
LoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/LoRA
QLoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QLoRA
DPOhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/DPO
QA-LoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/QA-LoRA
ReLoRAhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/ReLora
CPUhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/QLoRA-FineTuning
https://github.com/intel/ipex-llm#integration-with-community-libraries
HuggingFace transformershttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace
Standard PyTorch modelhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/PyTorch-Models
LangChainhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LangChain
LlamaIndexhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LlamaIndex
DeepSpeed-AutoTPhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/Deepspeed-AutoTP
Axolotlhttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Quickstart/axolotl_quickstart.md
HuggingFace PEFThttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/HF-PEFT
HuggingFace TRLhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/LLM-Finetuning/DPO
AutoGenhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Applications/autogen
ModeScopehttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/ModelScope-Models
Tutorialshttps://github.com/intel-analytics/ipex-llm-tutorial
https://github.com/intel/ipex-llm#api-doc
HuggingFace Transformers-style API (Auto Classes)https://github.com/intel/ipex-llm/blob/main/docs/mddocs/PythonAPI/transformers.md
API for arbitrary PyTorch Modelhttps://github.com/intel-analytics/ipex-llm/blob/main/docs/mddocs/PythonAPI/optimize.md
https://github.com/intel/ipex-llm#faq
FAQ & Trouble Shootinghttps://github.com/intel/ipex-llm/blob/main/docs/mddocs/Overview/FAQ/faq.md
https://github.com/intel/ipex-llm#verified-models
link1https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Native-Models
link2https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/vicuna
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/vicuna
link1https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Native-Models
link2https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/llama2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/llama2
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
C++ linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM/CPP_Examples
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/llama3
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/llama3
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
C++ linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM/CPP_Examples
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/llama3.1
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/llama3.1
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/llama3.2
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
C++ linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM/CPP_Examples
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/PyTorch-Models/Model/llama3.2-vision
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/chatglm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/chatglm2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/chatglm2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/chatglm3
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/chatglm3
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/glm4
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/glm4
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/glm-4v
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/glm-4v
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/glm-edge
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/glm-edge-v
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/mistral
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/mistral
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/mixtral
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/mixtral
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/falcon
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/falcon
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/mpt
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/mpt
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/dolly_v1
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/dolly-v1
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/dolly_v2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/dolly-v2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/replit
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/replit
link1https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Native-Models
link2https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/redpajama
link1https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Native-Models
link2https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/phoenix
link1https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/Native-Models
link2https://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/starcoder
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/starcoder
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/baichuan
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/baichuan
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/baichuan2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/baichuan2
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/internlm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/internlm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/internvl2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/qwen
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/qwen
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/qwen1.5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/qwen1.5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/qwen2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/qwen2
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
C++ linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM/CPP_Examples
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/qwen2.5
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
C++ linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM/CPP_Examples
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/qwen-vl
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/qwen-vl
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/qwen2-vl
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/qwen2-audio
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/aquila
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/aquila
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/aquila2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/aquila2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/moss
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/whisper
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/whisper
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/phi-1_5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/phi-1_5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/flan-t5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/flan-t5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/PyTorch-Models/Model/llava
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/PyTorch-Models/Model/llava
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/codellama
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/codellama
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/skywork
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/internlm-xcomposer
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/wizardcoder-python
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/codeshell
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/fuyu
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/distil-whisper
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/distil-whisper
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/yi
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/yi
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/bluelm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/bluelm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/PyTorch-Models/Model/mamba
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/PyTorch-Models/Model/mamba
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/solar
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/solar
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/phixtral
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/phixtral
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/internlm2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/internlm2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/rwkv4
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/rwkv5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/PyTorch-Models/Model/bark
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/PyTorch-Models/Model/bark
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/PyTorch-Models/Model/speech-t5
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/deepseek-moe
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/ziya
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/phi-2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/phi-2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/phi-3
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/phi-3
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/phi-3-vision
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/phi-3-vision
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/yuan2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/yuan2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/gemma
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/gemma
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/gemma2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/deciLM-7b
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/deciLM-7b
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/deepseek
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/deepseek
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/stablelm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/stablelm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/codegemma
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/codegemma
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/cohere
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/cohere
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/codegeex2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/codegeex2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/minicpm
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/minicpm
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM
C++ linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/LLM/CPP_Examples
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/minicpm3
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/MiniCPM-V
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/minicpm-v-2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/MiniCPM-V-2
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/MiniCPM-Llama3-V-2_5
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/Multimodal
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/CPU/HF-Transformers-AutoModels/Model/minicpm-v-2_6
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/MiniCPM-V-2_6
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/Multimodal
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/MiniCPM-o-2_6
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/janus-pro
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/LLM/moonlight
linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/GPU/HuggingFace/Multimodal/StableDiffusion
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/Embedding
Python linkhttps://github.com/intel/ipex-llm/blob/main/python/llm/example/NPU/HF-Transformers-AutoModels/Multimodal
https://github.com/intel/ipex-llm#get-support
Github Issuehttps://github.com/intel-analytics/ipex-llm/issues
GitHub Security Advisoryhttps://github.com/intel-analytics/ipex-llm/security/advisories
www.Intel.com/PerformanceIndexhttp://www.Intel.com/PerformanceIndex
https://github.com/intel/ipex-llm#user-content-fnref-1-e9586e16fae8e38ed833d96c96f33351
↩2https://github.com/intel/ipex-llm#user-content-fnref-1-2-e9586e16fae8e38ed833d96c96f33351
gpuhttps://github.com/topics/gpu
llmhttps://github.com/topics/llm
pytorchhttps://github.com/topics/pytorch
transformershttps://github.com/topics/transformers
Readmehttps://github.com/intel/ipex-llm#readme-ov-file
Apache-2.0 licensehttps://github.com/intel/ipex-llm#Apache-2.0-1-ov-file
Code of conducthttps://github.com/intel/ipex-llm#coc-ov-file
Contributinghttps://github.com/intel/ipex-llm#contributing-ov-file
Security policyhttps://github.com/intel/ipex-llm#security-ov-file
Activityhttps://github.com/intel/ipex-llm/activity
Custom propertieshttps://github.com/intel/ipex-llm/custom-properties
8.9k starshttps://github.com/intel/ipex-llm/stargazers
264 watchinghttps://github.com/intel/ipex-llm/watchers
1.4k forkshttps://github.com/intel/ipex-llm/forks
Report repositoryhttps://github.com/contact/report-content?content_url=https%3A%2F%2Fgithub.com%2Fintel%2Fipex-llm&report=intel+%28user%29
https://github.com
Termshttps://docs.github.com/site-policy/github-terms/github-terms-of-service
Privacyhttps://docs.github.com/site-policy/privacy-policies/github-privacy-statement
Securityhttps://github.com/security
Statushttps://www.githubstatus.com/
Communityhttps://github.community/
Docshttps://docs.github.com/
Contacthttps://support.github.com?tags=dotcom-footer

Viewport: width=device-width


URLs of crawlers that visited me.