mirror of https://github.com/ggerganov/llama.cpp.git synced 2025-04-14 10:36:07 +00:00

History

Georgi Gerganov e0dbec0bc6

llama : refactor llama_context, llama_kv_cache, llm_build_context (#12181 )

* llama : refactor llama_context, llama_kv_cache, llm_build_context

ggml-ci

* graph : don't mutate the KV cache during defrag

ggml-ci

* context : reduce virtuals + remove test function

ggml-ci

* context : move interface implementation to source file + factory

ggml-ci

* graph : move KV cache build functions to llama_context impl

ggml-ci

* graph : remove model reference from build_pooling

ggml-ci

* graph : remove llama_model reference

ggml-ci

* kv_cache : provide rope factors

ggml-ci

* graph : rework inputs to use only unique_ptr, remove attn input abstraction

ggml-ci

* context : remove llama_context_i abstraction

ggml-ci

* context : clean-up

ggml-ci

* graph : clean-up

ggml-ci

* llama : remove redundant keywords (struct, enum)

ggml-ci

* model : adapt gemma3

ggml-ci

* graph : restore same attention ops as on master

ggml-ci

* llama : remove TODO + fix indent

ggml-ci

2025-03-13 12:35:44 +02:00

llama.cpp.swift

llama : refactor llama_context, llama_kv_cache, llm_build_context (#12181 )

2025-03-13 12:35:44 +02:00

llama.swiftui

llama.swiftui : add "Done" dismiss button to help view (#11998 )

2025-02-22 06:33:29 +01:00

llama.swiftui.xcodeproj

llama : add xcframework build script (#11996 )

2025-03-05 06:30:31 +01:00

.gitignore

llama.swiftui : add bench functionality (#4483 )

2023-12-17 19:38:41 +02:00

README.md

llama.swiftui : fix xcframework dir in README [no ci] (#12353 )

2025-03-12 13:45:32 +01:00

README.md

llama.cpp/examples/llama.swiftui

Local inference of llama.cpp on an iPhone. This is a sample app that can be used as a starting point for more advanced projects.

For usage instructions and performance stats, check the following discussion: https://github.com/ggml-org/llama.cpp/discussions/4508

Building

First llama.cpp need to be built and a XCFramework needs to be created. This can be done by running the following script from the llama.cpp project root:

$ ./build-xcframework.sh

Open llama.swiftui.xcodeproj project in Xcode and you should be able to build and run the app on a simulator or a real device.

To use the framework with a different project, the XCFramework can be added to the project by adding build-apple/llama.xcframework by dragging and dropping it into the project navigator, or by manually selecting the framework in the "Frameworks, Libraries, and Embedded Content" section of the project settings.

Video demonstration:

https://github.com/bachittle/llama.cpp/assets/39804642/e290827a-4edb-4093-9642-2a5e399ec545