René's URL Explorer Experiment


Title: History for llama_cpp/server - hoangperry/llama-cpp-python · GitHub

Open Graph Title: History for llama_cpp/server - hoangperry/llama-cpp-python

X Title: History for llama_cpp/server - hoangperry/llama-cpp-python

Description: Python bindings for llama.cpp. Contribute to hoangperry/llama-cpp-python development by creating an account on GitHub.

Open Graph Description: Python bindings for llama.cpp. Contribute to hoangperry/llama-cpp-python development by creating an account on GitHub.

X Description: Python bindings for llama.cpp. Contribute to hoangperry/llama-cpp-python development by creating an account on GitHub.

Opengraph URL: https://github.com/hoangperry/llama-cpp-python

X: @github

direct link

Domain: github.com

route-pattern/:user_id/:repository/commits(/*name)
route-controllercommits
route-actionshow
fetch-noncev2:97079047-03c3-7da9-ef8f-767e3febfc6a
current-catalog-service-hashf3abb0cc802f3d7b95fc8762b94bdcb13bf39634c40c357301c4aa1d67a256fb
request-id93DE:C4AF2:1F88ABE:2B068B4:6A630500
html-safe-nonce35add552fd68de64288c69f2e5cba245b4c8a02887732008047fd66a3030094d
visitor-payloadeyJyZWZlcnJlciI6IiIsInJlcXVlc3RfaWQiOiI5M0RFOkM0QUYyOjFGODhBQkU6MkIwNjhCNDo2QTYzMDUwMCIsInZpc2l0b3JfaWQiOiI2MjQ0ODM2NDg1NDI4ODcyNDQ4IiwicmVnaW9uX2VkZ2UiOiJpYWQiLCJyZWdpb25fcmVuZGVyIjoiaWFkIn0=
visitor-hmacbbabe014fa18c76c09d9689169f08460e3d90560d6f3082122a6fc40b51fba5c
hovercard-subject-tagrepository:1241010171
github-keyboard-shortcutsrepository,commit-list,copilot
google-site-verificationApib7-x98H0j5cPqHWwSMm6dNU4GmODRoqxLiDzdx9I
octolytics-urlhttps://collector.github.com/github/collect
analytics-location///commits/show
fb:app_id1401488693436528
apple-itunes-appapp-id=1477376905, app-argument=https://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server
twitter:imagehttps://opengraph.githubassets.com/2ce9bfaa9b0d924abb79865f7f9bb92d5652d72204d0abf398d4d2362a261474/hoangperry/llama-cpp-python
twitter:cardsummary_large_image
og:imagehttps://opengraph.githubassets.com/2ce9bfaa9b0d924abb79865f7f9bb92d5652d72204d0abf398d4d2362a261474/hoangperry/llama-cpp-python
og:image:altPython bindings for llama.cpp. Contribute to hoangperry/llama-cpp-python development by creating an account on GitHub.
og:image:width1200
og:image:height600
og:site_nameGitHub
og:typeobject
hostnamegithub.com
expected-hostnamegithub.com
Noneb48e7bf0b1945d51f303503b057abf666ce74c3efbe107b413db2bb921c3f012
turbo-cache-controlno-cache
go-importgithub.com/hoangperry/llama-cpp-python git https://github.com/hoangperry/llama-cpp-python.git
octolytics-dimension-user_id37509456
octolytics-dimension-user_loginhoangperry
octolytics-dimension-repository_id1241010171
octolytics-dimension-repository_nwohoangperry/llama-cpp-python
octolytics-dimension-repository_publictrue
octolytics-dimension-repository_is_forktrue
octolytics-dimension-repository_parent_id617868717
octolytics-dimension-repository_parent_nwoabetlen/llama-cpp-python
octolytics-dimension-repository_network_root_id617868717
octolytics-dimension-repository_network_root_nwoabetlen/llama-cpp-python
turbo-body-classeslogged-out env-production page-responsive
disable-turbofalse
browser-stats-urlhttps://api.github.com/_private/browser/stats
browser-errors-urlhttps://api.github.com/_private/browser/errors
release9370f37042c29822d6ba7e3f670f0036a0155b21
ui-targetfull
theme-color#1e2327
color-schemelight dark

Links:

Skip to contenthttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server#start-of-content
https://github.com/
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2Fhoangperry%2Fllama-cpp-python%2Fcommits%2Fbatch-processing%2Fllama_cpp%2Fserver
GitHub CopilotWrite better code with AIhttps://github.com/features/copilot
GitHub Copilot appDirect agents from issue to mergehttps://github.com/features/ai/github-app
MCP RegistryNewIntegrate external toolshttps://github.com/mcp
ActionsAutomate any workflowhttps://github.com/features/actions
CodespacesInstant dev environmentshttps://github.com/features/codespaces
IssuesPlan and track workhttps://github.com/features/issues
Code ReviewManage code changeshttps://github.com/features/code-review
Code QualityEnforce quality at mergehttps://github.com/features/code-quality
GitHub Advanced SecurityFind and fix vulnerabilitieshttps://github.com/security/advanced-security
Code securitySecure your code as you buildhttps://github.com/security/advanced-security/code-security
Secret protectionStop leaks before they starthttps://github.com/security/advanced-security/secret-protection
Why GitHubhttps://github.com/why-github
Documentationhttps://docs.github.com
Bloghttps://github.blog
Changeloghttps://github.blog/changelog
Marketplacehttps://github.com/marketplace
View all featureshttps://github.com/features
Enterpriseshttps://github.com/enterprise
Small and medium teamshttps://github.com/team
Startupshttps://github.com/enterprise/startups
Nonprofitshttps://github.com/solutions/industry/nonprofits
App Modernizationhttps://github.com/solutions/use-case/app-modernization
DevSecOpshttps://github.com/solutions/use-case/devsecops
DevOpshttps://github.com/solutions/use-case/devops
CI/CDhttps://github.com/solutions/use-case/ci-cd
View all use caseshttps://github.com/solutions/use-case
Healthcarehttps://github.com/solutions/industry/healthcare
Financial serviceshttps://github.com/solutions/industry/financial-services
Manufacturinghttps://github.com/solutions/industry/manufacturing
Governmenthttps://github.com/solutions/industry/government
View all industrieshttps://github.com/solutions/industry
View all solutionshttps://github.com/solutions
AIhttps://github.com/resources/articles?topic=ai
Software Developmenthttps://github.com/resources/articles?topic=software-development
DevOpshttps://github.com/resources/articles?topic=devops
Securityhttps://github.com/resources/articles?topic=security
View all topicshttps://github.com/resources/articles
Customer storieshttps://github.com/customer-stories
Events & webinarshttps://github.com/resources/events
Ebooks & reportshttps://github.com/resources/whitepapers
Business insightshttps://github.com/solutions/executive-insights
GitHub Skillshttps://skills.github.com
Documentationhttps://docs.github.com
Customer supporthttps://support.github.com
Community forumhttps://github.com/orgs/community/discussions
Trust centerhttps://github.com/trust-center
Partnershttps://github.com/partners
View all resourceshttps://github.com/resources
GitHub SponsorsFund open source developershttps://github.com/open-source/sponsors
Security Labhttps://securitylab.github.com
Maintainer Communityhttps://maintainers.github.com
Acceleratorhttps://github.com/open-source/accelerator
GitHub Starshttps://stars.github.com
Archive Programhttps://archiveprogram.github.com
Topicshttps://github.com/topics
Trendinghttps://github.com/trending
Collectionshttps://github.com/collections
Enterprise platformAI-powered developer platformhttps://github.com/enterprise
GitHub Advanced SecurityEnterprise-grade security featureshttps://github.com/security/advanced-security
Copilot for BusinessEnterprise-grade AI featureshttps://github.com/features/copilot/copilot-business
Premium SupportEnterprise-grade 24/7 supporthttps://github.com/enterprise/premium-support
Pricinghttps://github.com/pricing
Search syntax tipshttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
documentationhttps://docs.github.com/search-github/github-code-search/understanding-github-code-search-syntax
Sign in https://github.com/login?return_to=https%3A%2F%2Fgithub.com%2Fhoangperry%2Fllama-cpp-python%2Fcommits%2Fbatch-processing%2Fllama_cpp%2Fserver
Sign up https://github.com/signup?ref_cta=Sign+up&ref_loc=header+logged+out&ref_page=%2F%3Cuser-name%3E%2F%3Crepo-name%3E%2Fcommits%2Fshow&source=header-repo&source_repo=hoangperry%2Fllama-cpp-python
Reloadhttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server
Reloadhttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server
Reloadhttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server
hoangperry https://github.com/hoangperry
llama-cpp-pythonhttps://github.com/hoangperry/llama-cpp-python
abetlen/llama-cpp-pythonhttps://github.com/abetlen/llama-cpp-python
Notifications https://github.com/login?return_to=%2Fhoangperry%2Fllama-cpp-python
Fork 0 https://github.com/login?return_to=%2Fhoangperry%2Fllama-cpp-python
Star 0 https://github.com/login?return_to=%2Fhoangperry%2Fllama-cpp-python
Code https://github.com/hoangperry/llama-cpp-python/tree/batch-processing
Pull requests 0 https://github.com/hoangperry/llama-cpp-python/pulls
Actions https://github.com/hoangperry/llama-cpp-python/actions
Projects https://github.com/hoangperry/llama-cpp-python/projects
Security and quality 0 https://github.com/hoangperry/llama-cpp-python/security
Insights https://github.com/hoangperry/llama-cpp-python/pulse
Code https://github.com/hoangperry/llama-cpp-python/tree/batch-processing
Pull requests https://github.com/hoangperry/llama-cpp-python/pulls
Actions https://github.com/hoangperry/llama-cpp-python/actions
Projects https://github.com/hoangperry/llama-cpp-python/projects
Security and quality https://github.com/hoangperry/llama-cpp-python/security
Insights https://github.com/hoangperry/llama-cpp-python/pulse
llama-cpp-pythonhttps://github.com/hoangperry/llama-cpp-python/commits
llama_cpphttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp
serverhttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server
batch-processinghttps://github.com/hoangperry/llama-cpp-python/tree/batch-processing
fix(cli): allow passing n_ctx=0 to openAI API server args to use model n_ctx_train field per #1015 (#1093)https://github.com/hoangperry/llama-cpp-python/commit/9c36688b33d32f77c501f4815047e4fc5581cf83
https://github.com/K-Mistele
K-Mistelehttps://github.com/hoangperry/llama-cpp-python/commits?author=K-Mistele
9c36688https://github.com/hoangperry/llama-cpp-python/commit/9c36688b33d32f77c501f4815047e4fc5581cf83
https://github.com/hoangperry/llama-cpp-python/blob/9c36688b33d32f77c501f4815047e4fc5581cf83/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/9c36688b33d32f77c501f4815047e4fc5581cf83
Support Accept text/event-stream in chat and completion endpoints, resolves #1083 (#1088)https://github.com/hoangperry/llama-cpp-python/commit/cfb7da98ed64feb882065222b17ca261356a7f59
cfb7da9https://github.com/hoangperry/llama-cpp-python/commit/cfb7da98ed64feb882065222b17ca261356a7f59
https://github.com/hoangperry/llama-cpp-python/blob/cfb7da98ed64feb882065222b17ca261356a7f59/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/cfb7da98ed64feb882065222b17ca261356a7f59
Add split_mode option. Closes #1085https://github.com/hoangperry/llama-cpp-python/commit/84615adbc6855c8384807c42f0130f9a1763f99d
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
84615adhttps://github.com/hoangperry/llama-cpp-python/commit/84615adbc6855c8384807c42f0130f9a1763f99d
https://github.com/hoangperry/llama-cpp-python/blob/84615adbc6855c8384807c42f0130f9a1763f99d/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/84615adbc6855c8384807c42f0130f9a1763f99d
Implement GGUF metadata KV overrides (#1011)https://github.com/hoangperry/llama-cpp-python/commit/76aafa61497ce4c9eda22c9892ae93ccb8f7d814
phiharrihttps://github.com/hoangperry/llama-cpp-python/commits?author=phiharri
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
76aafa6https://github.com/hoangperry/llama-cpp-python/commit/76aafa61497ce4c9eda22c9892ae93ccb8f7d814
https://github.com/hoangperry/llama-cpp-python/blob/76aafa61497ce4c9eda22c9892ae93ccb8f7d814/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/76aafa61497ce4c9eda22c9892ae93ccb8f7d814
docs: add server config docshttps://github.com/hoangperry/llama-cpp-python/commit/522aecb8689f1b98856c024bb47ad4640bb52072
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
522aecbhttps://github.com/hoangperry/llama-cpp-python/commit/522aecb8689f1b98856c024bb47ad4640bb52072
https://github.com/hoangperry/llama-cpp-python/blob/522aecb8689f1b98856c024bb47ad4640bb52072/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/522aecb8689f1b98856c024bb47ad4640bb52072
server: Support none defaulting to infinity for completions (#111)https://github.com/hoangperry/llama-cpp-python/commit/4b01a873efdef1c0f52dad87f11b9340ec8adc13
swghttps://github.com/hoangperry/llama-cpp-python/commits?author=swg
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
4b01a87https://github.com/hoangperry/llama-cpp-python/commit/4b01a873efdef1c0f52dad87f11b9340ec8adc13
https://github.com/hoangperry/llama-cpp-python/blob/4b01a873efdef1c0f52dad87f11b9340ec8adc13/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/4b01a873efdef1c0f52dad87f11b9340ec8adc13
[Feat] Multi model support (#931)https://github.com/hoangperry/llama-cpp-python/commit/12b7f2f4e9c68efa5555bacfaa49e44eb08efe60
D4ve-Rhttps://github.com/hoangperry/llama-cpp-python/commits?author=D4ve-R
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
12b7f2fhttps://github.com/hoangperry/llama-cpp-python/commit/12b7f2f4e9c68efa5555bacfaa49e44eb08efe60
https://github.com/hoangperry/llama-cpp-python/blob/12b7f2f4e9c68efa5555bacfaa49e44eb08efe60/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/12b7f2f4e9c68efa5555bacfaa49e44eb08efe60
Implement openai api compatible authentication (#1010)https://github.com/hoangperry/llama-cpp-python/commit/33cc623346d8c24ac4283f5b79d7154b821f62a7
https://github.com/docmeth02
docmeth02https://github.com/hoangperry/llama-cpp-python/commits?author=docmeth02
33cc623https://github.com/hoangperry/llama-cpp-python/commit/33cc623346d8c24ac4283f5b79d7154b821f62a7
https://github.com/hoangperry/llama-cpp-python/blob/33cc623346d8c24ac4283f5b79d7154b821f62a7/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/33cc623346d8c24ac4283f5b79d7154b821f62a7
Add offload_kqv option to llama and serverhttps://github.com/hoangperry/llama-cpp-python/commit/095c65000642a3cf73055d7428232fb18b73c6f3
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
095c650https://github.com/hoangperry/llama-cpp-python/commit/095c65000642a3cf73055d7428232fb18b73c6f3
https://github.com/hoangperry/llama-cpp-python/blob/095c65000642a3cf73055d7428232fb18b73c6f3/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/095c65000642a3cf73055d7428232fb18b73c6f3
Bugfix: Remove f16_kv, add offload_kqv field (#1019)https://github.com/hoangperry/llama-cpp-python/commit/62944df1425016018b5a137ebd0af5495eaa4894
https://github.com/brandonrobertz
brandonrobertzhttps://github.com/hoangperry/llama-cpp-python/commits?author=brandonrobertz
62944dfhttps://github.com/hoangperry/llama-cpp-python/commit/62944df1425016018b5a137ebd0af5495eaa4894
https://github.com/hoangperry/llama-cpp-python/blob/62944df1425016018b5a137ebd0af5495eaa4894/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/62944df1425016018b5a137ebd0af5495eaa4894
Add support for running the server with SSL (#994)https://github.com/hoangperry/llama-cpp-python/commit/8e44a32075de4aba2fc9877d4a2a34a0e7314c0d
https://github.com/rgerganov
rgerganovhttps://github.com/hoangperry/llama-cpp-python/commits?author=rgerganov
8e44a32https://github.com/hoangperry/llama-cpp-python/commit/8e44a32075de4aba2fc9877d4a2a34a0e7314c0d
https://github.com/hoangperry/llama-cpp-python/blob/8e44a32075de4aba2fc9877d4a2a34a0e7314c0d/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/8e44a32075de4aba2fc9877d4a2a34a0e7314c0d
docs: Update openapi endpoint nameshttps://github.com/hoangperry/llama-cpp-python/commit/1a7bf2037bba35b5b15340088694aa897d83fe36
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
1a7bf20https://github.com/hoangperry/llama-cpp-python/commit/1a7bf2037bba35b5b15340088694aa897d83fe36
https://github.com/hoangperry/llama-cpp-python/blob/1a7bf2037bba35b5b15340088694aa897d83fe36/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/1a7bf2037bba35b5b15340088694aa897d83fe36
Fix #569https://github.com/hoangperry/llama-cpp-python/commit/128dc4731fa846ead7e684a137ca57d8931b8899
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
128dc47https://github.com/hoangperry/llama-cpp-python/commit/128dc4731fa846ead7e684a137ca57d8931b8899
https://github.com/hoangperry/llama-cpp-python/blob/128dc4731fa846ead7e684a137ca57d8931b8899/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/128dc4731fa846ead7e684a137ca57d8931b8899
Formathttps://github.com/hoangperry/llama-cpp-python/commit/7a3f87846ba404ec573e10a022d12449c4dc7ab1
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
7a3f878https://github.com/hoangperry/llama-cpp-python/commit/7a3f87846ba404ec573e10a022d12449c4dc7ab1
https://github.com/hoangperry/llama-cpp-python/blob/7a3f87846ba404ec573e10a022d12449c4dc7ab1/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/7a3f87846ba404ec573e10a022d12449c4dc7ab1
Add support for logit_bias outside of server api. Closes #827https://github.com/hoangperry/llama-cpp-python/commit/07e47f55ba3a72e6022ebd12fb036373a7a7c4dd
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
07e47f5https://github.com/hoangperry/llama-cpp-python/commit/07e47f55ba3a72e6022ebd12fb036373a7a7c4dd
https://github.com/hoangperry/llama-cpp-python/blob/07e47f55ba3a72e6022ebd12fb036373a7a7c4dd/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/07e47f55ba3a72e6022ebd12fb036373a7a7c4dd
Added support for min_p (#921)https://github.com/hoangperry/llama-cpp-python/commit/b8438f70b58c1e4bbfb9581430ce4e032aaa3a8c
https://github.com/tk-master
tk-masterhttps://github.com/hoangperry/llama-cpp-python/commits?author=tk-master
b8438f7https://github.com/hoangperry/llama-cpp-python/commit/b8438f70b58c1e4bbfb9581430ce4e032aaa3a8c
https://github.com/hoangperry/llama-cpp-python/blob/b8438f70b58c1e4bbfb9581430ce4e032aaa3a8c/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/b8438f70b58c1e4bbfb9581430ce4e032aaa3a8c
Fix: default max_tokens matches openai api (16 for completion, max length for chat completion)https://github.com/hoangperry/llama-cpp-python/commit/e7962d2c733cbbeec5a37392c81f64185a9a39e8
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
e7962d2https://github.com/hoangperry/llama-cpp-python/commit/e7962d2c733cbbeec5a37392c81f64185a9a39e8
https://github.com/hoangperry/llama-cpp-python/blob/e7962d2c733cbbeec5a37392c81f64185a9a39e8/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/e7962d2c733cbbeec5a37392c81f64185a9a39e8
Fix destructor NoneType is not callable errorhttps://github.com/hoangperry/llama-cpp-python/commit/ca4cb8835165d2347b0cab926f89ee8060780b99
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
ca4cb88https://github.com/hoangperry/llama-cpp-python/commit/ca4cb8835165d2347b0cab926f89ee8060780b99
https://github.com/hoangperry/llama-cpp-python/blob/ca4cb8835165d2347b0cab926f89ee8060780b99/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/ca4cb8835165d2347b0cab926f89ee8060780b99
Add JSON mode support. Closes #881https://github.com/hoangperry/llama-cpp-python/commit/b30b9c338bf9af316d497ea501d39f5c246900db
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
b30b9c3https://github.com/hoangperry/llama-cpp-python/commit/b30b9c338bf9af316d497ea501d39f5c246900db
https://github.com/hoangperry/llama-cpp-python/blob/b30b9c338bf9af316d497ea501d39f5c246900db/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/b30b9c338bf9af316d497ea501d39f5c246900db
Add seed parameter support for completion and chat_completion requests. Closes #884https://github.com/hoangperry/llama-cpp-python/commit/86aeb9f3a14808575d2bb0076e6acb4a30907e6a
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
86aeb9fhttps://github.com/hoangperry/llama-cpp-python/commit/86aeb9f3a14808575d2bb0076e6acb4a30907e6a
https://github.com/hoangperry/llama-cpp-python/blob/86aeb9f3a14808575d2bb0076e6acb4a30907e6a/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/86aeb9f3a14808575d2bb0076e6acb4a30907e6a
Multimodal Support (Llava 1.5) (#821)https://github.com/hoangperry/llama-cpp-python/commit/aab74f0b2bd25cc2a9baeb743a2057dd5cada6e4
damian0815https://github.com/hoangperry/llama-cpp-python/commits?author=damian0815
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
aab74f0https://github.com/hoangperry/llama-cpp-python/commit/aab74f0b2bd25cc2a9baeb743a2057dd5cada6e4
https://github.com/hoangperry/llama-cpp-python/blob/aab74f0b2bd25cc2a9baeb743a2057dd5cada6e4/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/aab74f0b2bd25cc2a9baeb743a2057dd5cada6e4
Update llama.cpphttps://github.com/hoangperry/llama-cpp-python/commit/df9362eeea09a2f1e2649c3ad658673b6ad159d4
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
df9362ehttps://github.com/hoangperry/llama-cpp-python/commit/df9362eeea09a2f1e2649c3ad658673b6ad159d4
https://github.com/hoangperry/llama-cpp-python/blob/df9362eeea09a2f1e2649c3ad658673b6ad159d4/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/df9362eeea09a2f1e2649c3ad658673b6ad159d4
Add functionary support (#784)https://github.com/hoangperry/llama-cpp-python/commit/3af7b21ff1aec2ce4c2f8559e51de25907ed943d
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
3af7b21https://github.com/hoangperry/llama-cpp-python/commit/3af7b21ff1aec2ce4c2f8559e51de25907ed943d
https://github.com/hoangperry/llama-cpp-python/blob/3af7b21ff1aec2ce4c2f8559e51de25907ed943d/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/3af7b21ff1aec2ce4c2f8559e51de25907ed943d
Update llama.cpphttps://github.com/hoangperry/llama-cpp-python/commit/fa83cc5f9c8c0a7391c78afa5a4d25b0eb3e0090
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
fa83cc5https://github.com/hoangperry/llama-cpp-python/commit/fa83cc5f9c8c0a7391c78afa5a4d25b0eb3e0090
https://github.com/hoangperry/llama-cpp-python/blob/fa83cc5f9c8c0a7391c78afa5a4d25b0eb3e0090/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/fa83cc5f9c8c0a7391c78afa5a4d25b0eb3e0090
fix: tokenization of special characters: (#850)https://github.com/hoangperry/llama-cpp-python/commit/4d4e0f11e292984320e59f1dc1ecf9a0cc437a75
antoine-lizeehttps://github.com/hoangperry/llama-cpp-python/commits?author=antoine-lizee
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
4d4e0f1https://github.com/hoangperry/llama-cpp-python/commit/4d4e0f11e292984320e59f1dc1ecf9a0cc437a75
https://github.com/hoangperry/llama-cpp-python/blob/4d4e0f11e292984320e59f1dc1ecf9a0cc437a75/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/4d4e0f11e292984320e59f1dc1ecf9a0cc437a75
Iterate over tokens that should be biased rather than the entire vocabulary. (#851)https://github.com/hoangperry/llama-cpp-python/commit/3fc9147218ba503f0cfcd7f2f99b21c9a3e87fd0
https://github.com/zolastro
zolastrohttps://github.com/hoangperry/llama-cpp-python/commits?author=zolastro
3fc9147https://github.com/hoangperry/llama-cpp-python/commit/3fc9147218ba503f0cfcd7f2f99b21c9a3e87fd0
https://github.com/hoangperry/llama-cpp-python/blob/3fc9147218ba503f0cfcd7f2f99b21c9a3e87fd0/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/3fc9147218ba503f0cfcd7f2f99b21c9a3e87fd0
Pass-Through grammar parameter in web server. (#855) Closes #778https://github.com/hoangperry/llama-cpp-python/commit/5f8f369d1b675f3a6df3fb25cfa63995ec8957ba
https://github.com/dthuerck
dthuerckhttps://github.com/hoangperry/llama-cpp-python/commits?author=dthuerck
5f8f369https://github.com/hoangperry/llama-cpp-python/commit/5f8f369d1b675f3a6df3fb25cfa63995ec8957ba
https://github.com/hoangperry/llama-cpp-python/blob/5f8f369d1b675f3a6df3fb25cfa63995ec8957ba/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/5f8f369d1b675f3a6df3fb25cfa63995ec8957ba
update value check for n_gpu_layers field (#826)https://github.com/hoangperry/llama-cpp-python/commit/a315128d66ab4d93bd624f1c49343b4685fad4c7
https://github.com/hxy9243
hxy9243https://github.com/hoangperry/llama-cpp-python/commits?author=hxy9243
a315128https://github.com/hoangperry/llama-cpp-python/commit/a315128d66ab4d93bd624f1c49343b4685fad4c7
https://github.com/hoangperry/llama-cpp-python/blob/a315128d66ab4d93bd624f1c49343b4685fad4c7/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/a315128d66ab4d93bd624f1c49343b4685fad4c7
Print traceback on server errorhttps://github.com/hoangperry/llama-cpp-python/commit/d6a130a052db3a50975a719088a9226abfebb266
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
d6a130ahttps://github.com/hoangperry/llama-cpp-python/commit/d6a130a052db3a50975a719088a9226abfebb266
https://github.com/hoangperry/llama-cpp-python/blob/d6a130a052db3a50975a719088a9226abfebb266/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/d6a130a052db3a50975a719088a9226abfebb266
Log server exceptions to stdouthttps://github.com/hoangperry/llama-cpp-python/commit/5ef5280ef9830eb340b79c2896207f9ba3d5630d
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
5ef5280https://github.com/hoangperry/llama-cpp-python/commit/5ef5280ef9830eb340b79c2896207f9ba3d5630d
https://github.com/hoangperry/llama-cpp-python/blob/5ef5280ef9830eb340b79c2896207f9ba3d5630d/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/5ef5280ef9830eb340b79c2896207f9ba3d5630d
Update server paramshttps://github.com/hoangperry/llama-cpp-python/commit/d9bce17794d0dd6f7962d10aad768fedecf3ab89
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
d9bce17https://github.com/hoangperry/llama-cpp-python/commit/d9bce17794d0dd6f7962d10aad768fedecf3ab89
https://github.com/hoangperry/llama-cpp-python/blob/d9bce17794d0dd6f7962d10aad768fedecf3ab89/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/d9bce17794d0dd6f7962d10aad768fedecf3ab89
Adds openai-processing-ms response header (#748)https://github.com/hoangperry/llama-cpp-python/commit/3d5e5b1c0453345b771847539306dfef135062be
3d5e5b1https://github.com/hoangperry/llama-cpp-python/commit/3d5e5b1c0453345b771847539306dfef135062be
https://github.com/hoangperry/llama-cpp-python/blob/3d5e5b1c0453345b771847539306dfef135062be/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/3d5e5b1c0453345b771847539306dfef135062be
Remove confusing helpstring from server cli args. Closes #719https://github.com/hoangperry/llama-cpp-python/commit/b047b3034e28af43e30d864f652b3ac0e457e7f1
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
b047b30https://github.com/hoangperry/llama-cpp-python/commit/b047b3034e28af43e30d864f652b3ac0e457e7f1
https://github.com/hoangperry/llama-cpp-python/blob/b047b3034e28af43e30d864f652b3ac0e457e7f1/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/b047b3034e28af43e30d864f652b3ac0e457e7f1
Fix boolean env vars and cli argumentshttps://github.com/hoangperry/llama-cpp-python/commit/0449d29b9f940e437231a07b9d56550226558bac
https://github.com/abetlen
abetlenhttps://github.com/hoangperry/llama-cpp-python/commits?author=abetlen
0449d29https://github.com/hoangperry/llama-cpp-python/commit/0449d29b9f940e437231a07b9d56550226558bac
https://github.com/hoangperry/llama-cpp-python/blob/0449d29b9f940e437231a07b9d56550226558bac/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/0449d29b9f940e437231a07b9d56550226558bac
Update app.py (#705)https://github.com/hoangperry/llama-cpp-python/commit/58a6e42cc06d6cf6ad2bcb46e6e1627687d9695f
https://github.com/earonesty
earonestyhttps://github.com/hoangperry/llama-cpp-python/commits?author=earonesty
58a6e42https://github.com/hoangperry/llama-cpp-python/commit/58a6e42cc06d6cf6ad2bcb46e6e1627687d9695f
https://github.com/hoangperry/llama-cpp-python/blob/58a6e42cc06d6cf6ad2bcb46e6e1627687d9695f/llama_cpp/server
https://github.com/hoangperry/llama-cpp-python/tree/58a6e42cc06d6cf6ad2bcb46e6e1627687d9695f
Nexthttps://github.com/hoangperry/llama-cpp-python/commits/batch-processing/llama_cpp/server?after=850416ae825cb499d0aa6714e6688792c5c3e150+34
https://github.com
Termshttps://docs.github.com/site-policy/github-terms/github-terms-of-service
Privacyhttps://docs.github.com/site-policy/privacy-policies/github-privacy-statement
Securityhttps://github.com/security
Statushttps://www.githubstatus.com/
Communityhttps://github.community/
Docshttps://docs.github.com/
Contacthttps://support.github.com?tags=dotcom-footer

Viewport: width=device-width


URLs of crawlers that visited me.