cannot run fine-tuned gpt-oss model correctly
Maintainer thường phản hồi trong vòng 1 ngày
Chưa có ai nhận issue này.
Đánh giá
- Độ khó
- 4/5
- Thời gian dự kiến
- 3-5 ngày
- Mức phù hợp với người mới
- 25/100
- Loại issue
- Lỗi
- Độ rõ ràng
- Cần làm rõ
- Mức độ hoạt động
- Đình trệ
- Công nghệ
- python
- Lĩnh vực
- ai, backend-api-design, machine-learning
Hướng nghiên cứu
Issue báo cáo đầu ra không chính xác từ một mô hình GPT fine-tuned so với llama.cpp. Người dùng cung cấp đường dẫn mô hình và mã gọi. Bắt đầu bằng cách kiểm tra việc khởi tạo lớp LlamaCpp và phương thức invoke trong các bindings của llama-cpp-python. So sánh định dạng đầu ra và quá trình xử lý token với hành vi native của llama.cpp. Kiểm tra xem issue có liên quan đến 'harmony format' hoặc các chi tiết riêng trong việc tải mô hình đối với các mô hình fine-tuned hay không. Tái hiện lỗi bằng mô hình và prompt được cung cấp để xem đầu ra không khớp.
Do mô hình lập chỉ mục viết ra từ nội dung của issue.
Mô tả
Expected Behavior
Should produce the similar output format as llamacpp
Current Behavior
The output is wrong. Maybe related to the harmony format?
The current output of llamacpp-python:
Answer: weak
Llama.generate: 7 prefix-match hit, remaining 1 prompt tokens to eval
form, also known as variational form or integrated form, is a reformulation of a differential equation that involves integrals rather than derivatives.
Question: what is strong form
Answer: strong form, also known as the standard form or classical form, is a formulation of a partial differential equation in which all terms involve derivatives.
Yes, what is weak solution? how to set of a) to b)
how
for
theorem
How do you like
of
Questionary
To solve
Sure
What if
question: strong form
(what is the strong form
# Environment and Context
Please provide detailed information about your computer setup. This is important in case the issue is not reproducible except for under certain specific conditions.
* Physical (or virtual) hardware you are using, e.g. for Linux:
`$ lscpu`
Architecture: x86_64
CPU op-mode(s): 32-bit, 64-bit
Byte Order: Little Endian
Address sizes: 48 bits physical, 48 bits virtual
CPU(s): 96
On-line CPU(s) list: 0-95
Thread(s) per core: 2
Core(s) per socket: 24
Socket(s): 2
NUMA node(s): 2
Vendor ID: AuthenticAMD
CPU family: 25
Model: 1
Model name: AMD EPYC 7413 24-Core Processor
Stepping: 1
Frequency boost: enabled
CPU MHz: 1498.207
CPU max MHz: 2650.0000
CPU min MHz: 1500.0000
BogoMIPS: 5299.85
Virtualization: AMD-V
L1d cache: 1.5 MiB
L1i cache: 1.5 MiB
L2 cache: 24 MiB
L3 cache: 256 MiB
NUMA node0 CPU(s): 0-23,48-71
NUMA node1 CPU(s): 24-47,72-95
* Operating System, e.g. for Linux:
`$ uname -a`
Linux athena 5.4.0-176-generic #196-Ubuntu SMP Fri Mar 22 16:46:39 UTC 2024 x86_64 x86_64 x86_64 GNU/Linux
* SDK version, e.g. for Linux:
$ python3 --version
Python 3.10.18
$ make --version
GNU Make 4.2.1
$ g++ --version
g++ (Ubuntu 9.4.0-1ubuntu1~20.04.2) 9.4.0
# Failure Information (for bugs)
Just call the model, didn't produce the correct result. But llammacpp works well, give me the correct output.
# Steps to Reproduce
Please provide detailed steps for reproducing the issue. We are not sitting in front of your screen, so the more detail the better.
# Make sure the model path is correct for your system!
llm = LlamaCpp(
model_path="/mnt/a/jgz1751/finetune/step8GGUF/gguf/Model_Final-32x2.4B-Q8_0.gguf",
temperature=0.75,
max_tokens=2000,
top_p=1,
callback_manager=callback_manager,
verbose=True, # Verbose is required to pass to the callback manager
)
question = """
Question: what is weak form
"""
llm.invoke(question)
- Ngôn ngữ chính
- Python
- Star
- 10.6k
- Fork
- 1.5k
- Merge trung bình
- 3 giờ 57 phút
- Pull request đã merge (30 ngày)
- 4
Chuẩn bị môi trường
- Không có Dockerfile hay tệp Docker Compose
- Không có mẫu pull request
- Đọc hướng dẫn đóng góp
Bắt đầu từ đâu
- Đọc hết issue, rồi đọc hướng dẫn đóng góp của dự án.
- Bình luận trên issue rằng bạn sẽ nhận — tránh hai người làm cùng một việc.
- Fork repository và làm thay đổi trên một nhánh.
- Mở pull request có tham chiếu số hiệu của issue.
Issue khác của abetlen/llama-cpp-python
-
Seven llama_sampler_init_* bindings admit keyword arguments that the ctypes function object silently dropsCó thể đã có người làm @Belal0066 đã nhận 19 ngày trước. Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 88/100
abetlen/llama-cpp-python#2371 ·
Maintainer thường phản hồi trong vòng 1 ngày
-
uv add llama-cpp-python wheels fails for versions above 0.3.30Có thể đã có người làm Có pull request liên kết đang mở hoặc đã được merge. Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 65/100
abetlen/llama-cpp-python#2352 · 1 bình luận · 2 reaction ·
Maintainer thường phản hồi trong vòng 1 ngày
-
Docs: consolidate build-from-source and GPU backend guideCó thể đã có người làm Có pull request liên kết đang mở hoặc đã được merge. Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 75/100
abetlen/llama-cpp-python#2314 ·
Maintainer thường phản hồi trong vòng 1 ngày
-
Llama.embed() calls LlamaBatch.add_sequence with old 3-arg signature; missing logits_arrayCó thể làm lại được @lxcxjxhx đã nhận 96 ngày trước và không có pull request nào đang mở. Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 65/100
abetlen/llama-cpp-python#2211 · 2 bình luận ·
Maintainer thường phản hồi trong vòng 1 ngày
-
Llama() silently accepts and discards `embedding` kwarg; .embed() then raises confusinglyCó thể đã có người làm @Anai-Guo đã nhận 36 ngày trước. Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 65/100
abetlen/llama-cpp-python#2210 ·
Maintainer thường phản hồi trong vòng 1 ngày
Tất cả issue của abetlen/llama-cpp-python
Issue tương tự
-
Độ khó 1/5 Dưới một giờ Mức phù hợp với người mới 83/100
PedestrianDynamics/pyFDS-Evac#766 ·
Maintainer thường phản hồi trong vòng 1 ngày
-
Markdown tables render as literal text in 3 example files (missing blank line before header)Đang mở
Độ khó 1/5 1-3 giờ Mức phù hợp với người mới 91/100
alchaincyf/nuwa-skill#86 ·
-
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 76/100
Maintainer thường phản hồi trong vòng 2 ngày
-
Docs Needs Triage
Độ khó 1/5 Dưới một giờ Mức phù hợp với người mới 88/100
pandas-dev/pandas#71055 ·
Maintainer thường phản hồi trong vòng 1 ngày
-
[Bug]: graphify reads files that git's global ignore file hidesCó thể đã có người làm @smngvlkz đã nhận hôm nay. Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 72/100
Graphify-Labs/graphify#4335 · 1 bình luận ·
Maintainer thường phản hồi trong vòng 1 ngày