[Feature Request] End-to-end testing for OpenAI Agents SDK examples
@donald-pinckney is already working on this.
Since Sep 30, 2025.
Assessment
This issue has not been assessed yet.
Description
Is your feature request related to a problem? Please describe.
We need to add end-to-end tests to OpenAI Agents SDK samples to ensure that they remain stable.
Describe the solution you'd like
- Using the commands in the README, run each sample by launching the worker and the workflow runner program
- In some cases, additional resources, such as MCP servers or Postgres databases, may also need to run
- Some samples are interactive. Support this by sending commands to standard in.
- The core functionality should be conventional programmatic code.
- If the workflow runner exits with success, consider the run successful.
- Clean up all resources.
Additional context
- Optionally, verify: 1/ commands run match the README files (LLM or traditional search), 2/ Outputs match expected outputs (using LLMs).
- Dominant language
- Python
- Stars
- 367
- Forks
- 121
- Avg merge
- 3d 20h
- Merged PRs (30d)
- 11
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
More from temporalio/samples-python
-
enhancement
temporalio/samples-python#253 · 1 assignee ·
-
enhancement
Difficulty 3/5 1-2 days Newbie friendliness 45/100
temporalio/samples-python#231 · 1 comment ·
-
bug
Difficulty 4/5 3-5 days Newbie friendliness 35/100
temporalio/samples-python#192 · 2 comments ·
-
enhancement
Difficulty 3/5 1-2 days Newbie friendliness 45/100
temporalio/samples-python#191 ·
-
bug
Difficulty 1/5 Under an hour Newbie friendliness 48/100
temporalio/samples-python#184 ·
All issues in temporalio/samples-python
Similar issues
-
Difficulty 2/5 1-3 hours Newbie friendliness 78/100
syfoud/Simulated_Scepter#172 ·
-
A cancelled tests run makes the coverage comment workflow fail and reports it as a red check on main Openarea: ci bug perceived difficulty: 3
Difficulty 2/5 1-3 hours Newbie friendliness 78/100
Nitjsefnie-Harness-Commons/daedalus#921 · 1 comment ·
-
Difficulty 2/5 1-3 hours Newbie friendliness 86/100
EleutherAI/lm-evaluation-harness#4207 ·
-
Difficulty 1/5 Under an hour Newbie friendliness 92/100
-
Difficulty 2/5 1-3 hours Newbie friendliness 78/100
ClickHouse/clickhouse-connect#1057 ·