FineEnvs/repo2rlenv-swe-flow
repo2rlenv-swe-flow: Harbor dataset on Hugging Face with 100 tasks. Trace passing repository tests, select exercised source functions, remove their implementations and generate a reconstruction instruction. Restore the original functions as the reference solution and use the retained tests for…
Tasks
- Memoization is a common optimization technique that caches the results of expensive function calls and returns the cached result when the…
- When working with deeply nested data structures such as dictionaries and lists, it is often necessary to determine whether a particular…
- List manipulation utilities often require in-place element removal with flexible index targeting. A robust implementation must handle…
- Complex applications frequently work with deeply nested, heterogeneous data structures composed of lists, tuples, dicts, and sets.…
- A common utility need in Python development is the ability to construct dictionaries from two separate iterables — one providing keys and…
- Memoization decorators are a common pattern for caching the results of expensive function calls. A time-to-live (TTL) variant adds expiry…
- Many data processing and utility libraries require the ability to partition sequences into fixed-size chunks. This is a common pattern…
- In many software systems, certain initialization routines or expensive computations should be performed exactly once, regardless of how…
- Searching through sequences for the last occurrence of a value is a common operation in data processing. A utility that searches from a…
- When working with multiple related sequences of data, it is often necessary to sort all of them in tandem based on one or more of the…
- In data processing pipelines, it is common to need an inverse operation to cumulative aggregation. When values have been progressively…
- Dictionary manipulation is a common need in data processing pipelines. A utility that can produce a filtered view of a dictionary —…
- Many data processing and combinatorial tasks require the ability to slide a window of fixed size across a sequence while simultaneously…
- Platform tag generation is a core concern for package installation tooling. A correct implementation must enumerate compatible platform…
- Distribution packages carry metadata (version, name, dependencies, etc.) stored as RFC-2822-style email headers. A common need is to parse…
- When recording the origin of installed packages, URLs may contain authentication credentials (user names and passwords) embedded in the…
- When building developer tooling or debugging utilities, it is often necessary to produce human-readable representations of function calls…
- Many data processing pipelines need to partition a potentially large or infinite stream of items into fixed-size groups for downstream…
- When working with iterable data structures, it is often necessary to extract a single, unambiguous value from a collection. A utility that…
- When distributing compiled Python packages for macOS, installers and package managers must determine which binary formats a given platform…
- When working with deeply nested data structures such as dictionaries, lists, or combinations thereof, a common need arises to safely…
- SQL parsing libraries often need to handle identifier strings that may be wrapped in various quoting styles depending on the SQL dialect…
- Working with deeply nested data structures (dictionaries, lists, or combinations thereof) is a common need in data processing. Providing a…
- When working with multiple dictionaries, a common need is to align their values by shared keys and process those values together. A…
- Many utility libraries need a unified way to iterate over the "items" of a collection regardless of whether the collection is a mapping…
- In data processing pipelines, it is common to work with iterables that may contain elements causing exceptions when processed by a given…
- SQL formatting tools require a robust mechanism for accepting, validating, and normalizing user-supplied configuration options before…
- Command-line interfaces for SQL processing tools must handle a variety of argument combinations, file inputs, and error conditions…
- Work in /workspace . Submit your fix in the existing Python source files under more itertools/more.py , more itertools/recipes.py .…
- Iterator utility libraries often need to provide higher-level abstractions over standard iteration patterns. Two common needs are…
- Python packaging ecosystems rely on wheel tag strings to describe compatibility between a distribution and a target environment. Parsing…
- When working with multiple dictionaries, a common need is to find the intersection of their keys and retrieve corresponding values from…
- Memoization decorators are a common tool for improving the performance of functions by caching previously computed results. Different…
- Managing log files, backups, or any periodically overwritten files often requires a rotation mechanism that preserves a configurable…
- Searching sequences for element occurrences is a fundamental operation in data processing. While forward index searches are common,…
- Numerical iteration over floating-point ranges is a common requirement in scientific and data-processing applications. Unlike…
- List manipulation is a fundamental operation in many programming tasks. A common need is to append one or more items to an existing list…
- Wheel filenames encode structured metadata — project name, version, optional build tag, and compatibility tags — in a standardized format.…
- Python packaging tools need to identify which built distributions (wheels, sdists, etc.) are compatible with a given Python interpreter…
- When working with iterable data structures, a common need is to retrieve the first element efficiently and safely. This includes handling…
- When working with higher-order functions, decorators, or dynamic code generation, it is often necessary to produce an independent copy of…
- Many data processing workflows require grouping, transforming, and aggregating items from a sequence based on custom key and value logic.…
- When working with nested data structures such as dictionaries and lists, it is often necessary to remove a value at a specific nested…
- When working with deeply nested data structures such as dictionaries, lists, or combinations thereof, it is often necessary to verify…
- Collection manipulation and transformation are common tasks in functional programming utilities. A library needs to provide consistent,…
- Caching libraries commonly require a mechanism to produce stable, hashable keys from arbitrary positional and keyword arguments. Such keys…
- Managing object lifetimes and diagnosing memory leaks requires a reliable mechanism for enumerating all live instances of a given type at…
- When working with nested data structures such as dictionaries and lists, it is common to need a way to remove a specific element at an…
- Combinatorial mathematics frequently requires computing the lexicographic position (index) of a specific combination within the ordered…
- When working with sequences and iterables, it is often useful to enumerate all contiguous subsequences (substrings) of varying lengths. A…
- Python packaging tools need to determine which wheel files or source distributions are compatible with a given Python interpreter and…
- Many data processing pipelines require the ability to combine or align sequences with positional offsets, enabling comparisons or…
- When working with multiple dictionaries, a common need is to combine or correlate values that share the same key across all dictionaries.…
- When distributing or installing software on Linux systems, it is often necessary to determine which C runtime library an executable is…
- Memoization decorators are a common pattern for improving the performance of functions by caching previously computed results. Different…
- When working with sequences or collections, it is often useful to enumerate all contiguous subsequences along with their positional…
- Array manipulation utilities often require extracting contiguous sub-sequences from a list or sequence. A general-purpose slice utility…
- When working with higher-order functions, decorators, or dynamic code generation, it is sometimes necessary to produce an independent copy…
- When building developer tools, logging systems, or debugging utilities, it is often necessary to produce human-readable representations of…
- Caching systems often need to distinguish between arguments that compare as equal but have different types (e.g., the integer 1 and the…
- When processing sequences of data, it is often necessary to extend them to a desired length or ensure they conform to a specific size…
- Managing object lifecycles in long-running Python applications can be challenging. Developers sometimes need to enumerate all live…
- Set-union operations over collections frequently require custom notions of equality—for example, treating elements as identical when they…
- Searching sequences for the last occurrence of a value is a common operation in data processing and array manipulation utilities. A robust…
- Many applications require the ability to partition sequences into sub-sequences based on dynamic, user-supplied conditions. A…
- When working with multiple sequences of varying lengths, it is often desirable to merge them into a single sequence such that the elements…
- Sequence manipulation is a common requirement in utility libraries. One fundamental operation is reversing the order of elements in a…
- Caching mechanisms for methods often require a stable, hashable key derived from the arguments passed to the method. When methods are…
- Object-oriented systems frequently need to memoize the results of expensive method calls to avoid redundant computation. A general-purpose…
- When working with deeply nested data structures such as combinations of dictionaries and lists, accessing values at arbitrary depths…
- Itertools and iterator utility libraries often need robust, well-defined functions for retrieving specific elements from arbitrary…
- Package and distribution name handling in Python ecosystems requires consistent normalization so that names differing only in case,…
- Wheel distribution filenames embed compatibility tags that describe the Python interpreter, ABI, and platform a wheel supports. These tags…
- Format strings in Python can contain positional placeholders that are either anonymous (no explicit index) or explicitly numbered.…
- When working with deeply nested data structures (such as dictionaries, lists, or combinations thereof), it is common to need safe…
- When working with deeply nested data structures such as dictionaries, lists, or other indexable objects, accessing values at arbitrary…
- Memoization utilities are commonly used to cache the results of expensive function calls and return the cached result when the same inputs…
- Many data processing workflows require dividing a flat sequence of elements into smaller, evenly-sized groups for batch processing,…
- Format string utilities are commonly needed when building systems that process or transform Python-style format strings. One recurring…
- Distribution metadata for Python packages is commonly stored in an email-header format. Parsing such metadata into structured Python…
- In many applications, certain initialization routines or expensive computations should only run once regardless of how many times they are…
- Package management ecosystems require reliable utilities for validating project names against normalization rules. A normalized name is a…
- Processing deeply nested, heterogeneous data structures is a common need in data transformation pipelines. A general-purpose recursive…
- File system traversal utilities are commonly needed to locate files matching specific naming conventions across directory trees. Such…
- When working with heterogeneous collections, it is often necessary to produce an empty counterpart of a given collection while preserving…
- SQL tooling frequently needs to parse multi-statement SQL text and return each statement as a discrete string. This is useful for…
- Partitioning a collection into non-overlapping, non-empty subsets is a classical combinatorial problem with applications in grouping,…
- Set-like union operations on collections are a common need in data processing pipelines. When the definition of "equality" between…
- When working with combinatorial sequences, it is often necessary to efficiently retrieve a specific element at a given index without…
- Caching mechanisms for instance methods require a way to generate cache keys that are independent of the object instance ( self ). This…
- Many data processing workflows require dividing a flat sequence into fixed-size groups for batch processing, pagination, or parallel…
- File management systems often need to maintain a rolling set of backup copies of a file, cycling older versions out as new ones are…
- List manipulation utilities often require removing elements at arbitrary positions and returning the removed value, while also mutating…
- When working with nested data structures such as dicts and lists, it is often necessary to produce updated copies without mutating the…
- Combinatorial indexing is a common need in discrete mathematics and algorithm design. Given a pool of elements and a selection rule…
- Command-line SQL formatting tools need a well-structured argument parser to accept user input and route it to the appropriate formatting…
- When working with collections of mixed values, it is common to need a way to filter out logically false or empty values, retaining only…
- Work in /workspace . Submit your fix in the existing Python source files under more itertools/more.py , more itertools/recipes.py .…
- In many data processing workflows, it is necessary to partition a sequence of items into multiple sub-groups of specified lengths. A…
- Sorting collections is a fundamental operation in data manipulation libraries. A flexible sort utility must handle multiple sorting…