diff --git a/README.md b/README.md index 214bd7e..fa05bee 100644 --- a/README.md +++ b/README.md @@ -49,9 +49,11 @@ you want to erase the evidence and start fresh. - **Inference:** llama.cpp compiled from source (pinned tag `b10333`) through the NDK. Pulled automatically at build time via CMake `FetchContent` — no submodule to init. -- **Model:** `Qwen3.5-2B-Uncensored-HauhauCS-Aggressive-Q6_K.gguf` (~1.5 GB) is - **downloaded on first launch** from Hugging Face into the app's private storage. - It is *not* bundled in the APK. (Swap the quant in `ModelDownloader.kt`.) +- **Model:** + [`Qwen3.5-2B-Uncensored-HauhauCS-Aggressive`](https://huggingface.co/HauhauCS/Qwen3.5-2B-Uncensored-HauhauCS-Aggressive) + in Q6_K format (~1.5 GB) is **downloaded on first launch** from Hugging Face + into the app's private storage. It is *not* bundled in the APK. (Swap the + quant in `ModelDownloader.kt`.) - **Persona:** set via `SYSTEM_PROMPT` in `MainActivity.kt`; sampling is a little hot (temp 0.9) for playful answers. The model is a Qwen3 "thinking" model, so the prompt disables reasoning (empty `` prefill + `/no_think`) to keep @@ -99,6 +101,24 @@ manager, and tap it. First launch downloads the ~1.5 GB model over Wi-Fi. | Model URL / filename | `ModelDownloader.kt` | | Generation params (ctx, temp, tokens) | `MainActivity.kt` + `llama-jni.cpp` | +## Credits + +tsjetpiti is small because it stands on the shoulders of some decidedly +not-small projects: + +- [Qwen3.5-2B](https://huggingface.co/Qwen/Qwen3.5-2B) by the Qwen team is the + upstream base model. +- [Qwen3.5-2B-Uncensored-HauhauCS-Aggressive](https://huggingface.co/HauhauCS/Qwen3.5-2B-Uncensored-HauhauCS-Aggressive) + by [HauhauCS](https://huggingface.co/HauhauCS) is the model variant and GGUF + quantization used by the app. +- [llama.cpp](https://github.com/ggml-org/llama.cpp) by the ggml-org community + provides the on-device inference engine and GGUF runtime. +- [Android](https://developer.android.com/) and + [Kotlin](https://kotlinlang.org/) provide the application platform and native + app layer. +- [Gradle](https://gradle.org/) and [CMake](https://cmake.org/) power the Kotlin + and C++ build. + ## TODO / notes - **Branding:** launcher icon (adaptive + legacy densities) is generated from