Apex
Back to home

Training Data

Apex — Training Data Summary

Version 1.0 · Published 1 October 2026

This summary describes the content Callstack used to post-train Apex. It follows the structure of the European Commission's Template for the Public Summary of Training Content for general-purpose AI models (July 2025).

Scope. Apex is a modification of a general-purpose AI model released by Qwen (Alibaba Group). This summary covers only the content Callstack added in post-training. The content used to train the base model is described in the documentation published by Qwen.

Status. On the basis of the compute used for our modification, measured against the indicative criterion in the Commission's Guidelines on the scope of obligations for providers of general-purpose AI models (one third of the training compute of the original model), Callstack is not the provider of a general-purpose AI model in respect of Apex. We publish this summary voluntarily, because we believe customers and rightsholders should be able to see what went into the model they use.


1. General information

1.1 Provider

Name Callstack.io sp. z o.o.
Address ul. Prosta 36, 53-508 Wrocław, Poland
Contact apex@callstack.com

1.2 Model

Model name Apex
Versions covered apex-20261001, served under the apex alias
Base model A general-purpose model released by Qwen (Alibaba Group), used under licence from Qwen
Nature of modification Domain post-training for software engineering with React Native and React
Date placed on the market 1 October 2026

1.3 Training content — size and characteristics

Modality Used in post-training Size
Text, including source code Yes Less than 1 billion tokens
Images No —
Audio No —
Video No —

Characteristics. Source code in JavaScript, TypeScript, Kotlin, Java, Swift, Objective-C, Objective-C++ and C++, together with technical documentation, READMEs and code comments. Natural-language content is predominantly in English.

Latest date of data collection: September 2026.


2. List of data sources

2.1 Publicly available datasets

None. Callstack did not use any pre-compiled public dataset.

2.2 Private datasets licensed from third parties

None.

2.3 Content collected from online sources

Callstack obtained source code and documentation directly from public open-source repositories and package registries. No general web crawling was carried out.

Selection. The repositories were selected deliberately rather than crawled: the React and React Native core repositories, and the most widely used open-source libraries and packages in the React Native ecosystem, chosen by adoption and download volume.

Principal repositories:

  • React — github.com/react/react — MIT License
  • React Native — github.com/react/react-native — MIT License

Libraries and packages. Open-source React Native libraries and packages, each distributed under the MIT License. The complete list of repositories, with their licences, is available on request from apex@callstack.com.

Domains from which content was obtained:

Domain Content
github.com Source repositories
registry.npmjs.org Published package contents

Content obtained: source code files, documentation files, READMEs, licence files and in-repository examples. Commit histories, issue trackers, pull request discussions and other user-generated discussion content were not included.

2.4 User data

None. Callstack processes all customer API traffic under zero data retention: prompts and outputs are never stored, and so cannot be used for training.

2.5 Synthetic data

Instruction and task data derived from the sources in Section 2.3 and 2.6 was generated to teach the model to apply that material to development tasks. This synthetic data was generated using the base model and Callstack's own tooling, and was not produced using any third-party model whose terms prohibit its use for model training.

2.6 Other sources

Callstack internal documentation. Technical documents written by Callstack describing approaches to greenfield and brownfield React Native development and to migrations between frameworks, architectures and versions. Callstack owns this content. It was reviewed before use to exclude client-confidential information and personal data.


3. Relevant data processing aspects

3.1 Respect of reservations of rights from text and data mining

All third-party content used in post-training was obtained from repositories whose licences — in every case the MIT License — expressly permit use, copying, modification and the creation of derivative works. Callstack relied on those licences as the basis for use, and did not rely on the text and data mining exceptions in Articles 3 and 4 of Directive (EU) 2019/790.

Callstack did not include content from any repository whose licence does not permit such use, or which is published without a licence. No content was collected by crawling; accordingly, no robots.txt or other machine-readable opt-out mechanism was relevant to collection. Callstack maintains a policy for compliance with Union copyright law, available on request.

Rightsholders who believe their content has been used contrary to their rights may contact apex@callstack.com.

3.2 Removal of illegal content

The post-training content consisted of source code and technical documentation from established, maintained open-source projects and Callstack's own documents. The sources were selected individually, which by design excludes the categories of illegal content — including child sexual abuse material and terrorist content — that the removal measures in the Template are aimed at. Content was additionally screened for credentials, secrets and personal data before use.

3.3 Other information

Attribution and licence notices. The MIT License requires that its copyright and permission notice be included in copies or substantial portions of the licensed software. Apex is trained to generate new code rather than reproduce its training material, but it may occasionally output passages closely resembling code from the sources above. Users are responsible for reviewing Outputs and complying with any licence obligations that apply to material they incorporate.

Updates. This summary will be updated whenever Apex is further trained on content that changes the information above.