Hacktoberfest 2026: the issues maintainers tagged for October, open and beginner-friendly. Browse Hacktoberfest issues

Test vision encoder only frames 32 out of memory on 16GB GPU

Open
#97 6 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Assessment

Difficulty
4/5
Estimated time
3-5 days
Newbie friendliness
35/100
Issue type
Bug
Clarity
Needs clarification
Activity status
Stale
Tech stack
python, pytorch

Research direction

No source file or test is identified. Start by reproducing the vision-encoder run on a 16GB GPU with eight frames, then inspect the reported torch.Size outputs and memory use. Done means the frame-limit or out-of-memory behavior is understood and the issue has a verified resolution.

Written by the indexing model from the issue text.

Description

torch.Size([1, 8192, 1024])
torch.Size([1, 1024])

this is 8frame output last hiddent state, the tokens is extremly large.

Dominant language
Python
Stars
402
Forks
20
PR merge metrics
No merged PRs in 30d

Getting set up

  • Ships a Dockerfile or Docker Compose file
  • No pull request template
  • No contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

More from EvolvingLMMs-Lab/OneVision-Encoder

All issues in EvolvingLMMs-Lab/OneVision-Encoder

Similar issues

More Python issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.