Fetching the latest programs, projects, and workspace data.
Find open source projects actively accepting contributors. Search repositories, filter by program milestones, difficulty tags, or tech stack.
Use our Orbit AI Matcher to find out! Get instant matching scores based on your developer skills, preferred frameworks, and contribution experience.
Convert your selected open-source project into a winning GSoC, LFX, or Outreachy application using Proposal Studio.
MetaBrainz contains multiple sub-projects which sends out standalone notifications. This project aims to centralize those by developing a shared notification system within metabrainz-org, enabling all sub-projects to deliver user notifications through this notification system. Expected Outcome: A functional notifications system with relevant API endpoint.
<p>Thinking everyone got a message, but for some it didn't arrive – that is bad. This project will improve Matrix to better handle cases in which bridges couldn't deliver what they were asked for. It does so by adding new failure modes for bridges which can inform the user that something just isn't quite right.</p>
<p>This project adds a new feature to Worldbrain's MEMEX a web extension. The feature enables users to add Custom Lists. A List is a collection of pages that the user would like to group together. This project implements the UI/UX for adding and deleting custom lists as well as adding/deleting Pages to the list. As well all related functions and modules.</p>
<p>Over 100,000 people have already been genotype through DTC genetic testing however, these data live in siloes, often inaccessible to the public. {Bastian Greshake, Philipp E. Bayer, Helge Rausch, Julia Reda, openSNP: A Crowdsourced Web Resource for Personal Genomics, 2014}. openSNP seeks to solve this problem by allowing users to upload their results, where they'll be freely-available for all.</p> <p>In addition to these genetic data, openSNP's broader goal relates to all types of quantified-self (QS) information as this is of interest to genetic researchers.</p> <p>So far, openSNP allows people to link their Fitbit accounts, thereby donating their activity, sleep, and weight data to the public domain. By now, there are many more tracking devices (HTDs) around, and many of them offer their APIs. By exploiting these services, openSNP can dramatically increase the amount of QS data available in the public domain, providing a means for academic researchers and scientists to find new associations. These associations are easier to find in openSNP than similar services since a user's genetic test results, their tagged phenotypes, and HTD data are all linked to a single person.</p>
<p>3Scale’s API Management service allows us to manage APIs with features like authorize, rate-limit, and monetize, etc.In this project, we are going to implement an in-proxy cache as an Envoy filter extension to add necessary smarts to avoid calling SM API for each request requiring authorization. The target benchmark to beat would be “1 auth and rep (using authrep.xml endpoint)” per request. The filter will be written in Rust and then later “translated” into WASM code packaged as WASM Module which is loaded on WASM VMs spun up by a runtime (V8, WAVM, or Wasmtime) embedded inside the Envoy. Many advantages come with using the WASM module instead of the native C++ extension like dynamically loading the filter extension, abstraction from the underlying CPU and programming language used, and being safe from crashing the proxy when a developer makes a mistake. Even though the module is sandboxed, it has the capability for I/O as described in the ABI Specification which empowers us with access to shared data storage (key-value based) and a message queue used in this project.</p>
<p>The goal of this project would be to develop a web based user interface to visualise the variant prediction data from VEP plugin using the neXtProt tools to make it easily accessible to the users. The interface would be designed in such a way that it is performant, scalable and can easily handle large amounts of data that is computed by the VEP plugin API.</p>
<p>There is an underlying issue with scale degree block as we have it now. The current block does not perform the function that musicians expect when they think of scale degree. Instead the block functions in a way that we can specify a key/mode of a pitch length, input a number and result is a pitch in the chosen key/mode.<br> The current block has its utility in programming and we aim to keep it as such with a modified name. A new block for scale degree needs to be added with the desired functionality</p>
<p>MusicBrainz for Android was first created in 2010-11. In the year 2019,the app went through a significant update as part of GsoC’19 .The complete codebase was shifted to adhere to the new Android Ecosystem.The relevant updates had been made.The core of the app is now clean and in a much better condition. However,basic feaures like the Tagging facility is yet to be implemented.Also,the app’s UI must be completely redesigned and implemented as it needs to have a responsive layout in order to stay relevant in the current Android Environment and to keep up to the current Android Users’ need.</p> <p>My Project is aimed at:</p> <p>1 Fixing existing design issues and integrating the android debug bridge.</p> <p>2 Completing the Tag feature in all aspects and replace the existing library.</p> <p>3 Creating and integrating a responsive UI for users for which redesigning of some existing features and adding new features are proposed.This will be done in 2 steps. In the first one, I will add backend code for them and in the second one, the necessary GUI will be added.</p>
Using HTML CSS and JavaScript to-, 1. Maintain the Music Block v3 and fixing bugs related to Styling, Browser specific, improving the overall layout of the website and solving High priority issues . 2. Adding new functionalities like Recording feature , Full screen mode ,and touch scrolling
Crypt4GH is a file format developed by GA4GH that keeps genomic data encrypted at rest and in transit. Currently, implementations of the TES API do not support processing Crypt4GH files. I propose an implementation of the TES API that handles Crypt4GH files through the use of middleware that alters the initial TES request in order decrypt Crypt4GH files before processing. First, I will create a "TES playground" that is able to run a simple task. Then, I will add to the functionality of this playground in order to demonstrate an implementation of the TES API that can handle a Crypt4GH-encrypted file as an input. This project will hopefully serve as a proof-of-concept to inform future standard-setting works.
<p>We all know that nowadays a website needs a responsive layout, by the fact that there are a lot of different devices sizes, like smartphones, tablets and desktops. In addition to this, the project organization is another important thing: letting the code more readable and easier to maintain helps the developer to modify the code when something goes wrong. Another important thing that we can't ignore is the website performance, this one depends of the website weight, but, why performance is important? Because nowadays people wants instant information, since they have a hurry life. This project has as main goal to give these three important things to the OpenSNP website.</p>
<p>The goal of this project is to develop a tool which will allow to visualise the events Directed Acyclic Graph data structure which describes the conversation history in a room. It will be a real-time visualisation of the DAG of a given Matrix room, as seen from the perspective of one or more HomeServers (HSes).</p> <p>This tool will be useful for debugging or administration of Matrix HSes by making people able to easily see how the federation process works.</p>
<p>The goal of this proposal would be to implement the <a href="http://samtools.github.io/hts-specs/htsget.html" target="_blank">htsget protocol</a> in pure Rust using the <a href="https://github.com/zaeleus/noodles" target="_blank">noodles crate</a>. It also has the objective of making a serverless implementation able to run as a service under the <a href="https://github.com/brainstorm/s3-rust-htslib-bam" target="_blank">AWS Lambda Payload</a></p>
<p>This proposal is aimed at developing infrastructure for importing data into the BookBrainz database from third party sources, while at the same time ensuring that the quality of the data is maintained.</p>
<p>The main goal of this project is to reexamine the assumptions underlying the segregation of rhythm from pitch in these widgets and to design and implement a more unified experience. This can be done by improving the already present widgets inside music blocks as it will not cause any changes in the usage of any other section of music blocks and will not confuse the existing users . Currently in music blocks a user has to generate a rhythm by using rhythm maker widget and then import it inside another widget like phrase maker. I would like to streamline this process by including the functionality of editing rhythms inside phrase maker and in the case of musical keyboard the user can not even import rhythms so I will like to propose a function where the widget will be able to calculate the duration for which a key is pressed and the functionality of playing keyboard with the help of keyboard.</p>
<p>Reactome currently uses a REST-based API that allows end-users to obtain specific data from a set of predefined end-points. To provide better flexibility and allow users to query whatever data they need, this project provides a GraphQL interface to Reactome data fetched from a Neo4j database.</p> <p>There are several advantages of using GraphQL over RESTful API implementation:</p> <ol> <li>Request and retrieve exact data in a predictable manner.</li> <li>Through a single request, the ability to retrieve data from multiple sources.</li> <li>Since GraphQL is strongly-typed, it allows the user to know the availability and the state of data at any given point of time (through subscriptions).</li> </ol>
Milvus currently lacks a Debezium source connector, meaning data can flow into Milvus (via sink connectors) but not out of it into the broader data ecosystem. This critical gap forces users into manual workarounds for data replication, auditing, and prevents the creation of hybrid pipelines that join vector search results with relational data. The plan is to implement a Hybrid strategy focused exclusively on the future-proof Milvus 2.6+ StreamingNode architecture. This approach will leverage two separate sources to guarantee a complete and ordered change stream: Woodpecker Streaming Service (WAL): Used as the only source for all Data Manipulation Language (DML) events (inserts and deletes). Etcd Watch: Used as the cleanest source for all Data Definition Language (DDL) events (schema changes), which simplifies parsing and aligns with Debezium's schema history mechanism.
The objective of this project is to provide alias support in concerto language . For example import statement like "import {document as d} from Library" to be supported by the end of this project. This would allow us to import two models with same name in a single file by aliasing them to different names which is currently not supported in the language. Doing this would require changes in both the concerto-parser and concerto-runtime . The project requires compiler design and programming language knowledge. The stack used in project is NodeJs and written in typescript. The concerto-parse is written in PEGJS.
Build WebGL spatial telemetry panels in Grafana for visualizing physical data center rack thermal sensor metrics.
BookBrainz still has a relatively small community and contains less entities than other comparable databases. Therefore we want to provide a way to import available collections of library records into the database while still ensuring that they meet BookBrainz' high data quality standards. From a previous GSoC project, the database schema already contains additional tables set up for that purpose, where the imports will await a user's approval before becoming a fully accepted entity in the database. The project will require processing very large data dumps (e.g. MARC records or JSON files) in a robust way and transforming entities from one database schema to the BookBrainz schema. Additionally the whole process should be repeatable without creating duplicate entries.
Music Blocks v4 has a modular architecture with a block editor, compiler, and execution engine built separately but not yet connected. This project integrates the Masonry visual block editor into the main application, implements block snapping and connection logic so users can build structured programs, and wires the full execution pipeline (blocks to AST to IR to output) so programs produce real music and graphics. Core deliverables: (1) Masonry module integration with palette and workspace (2) snap-based block connections with visual feedback enforcing programming semantics (3) end-to-end execution connecting blocks to the Painter and Singer modules via a real function registry. Stretch goals include inline block editing and workspace zoom/pan/undo.
Modern AI systems relying on Retrieval-Augmented Generation (RAG) often suffer from stale data due to batch-based updates. This project, PyDebeziumAI, aims to solve this by enabling real-time synchronization between databases and LLM pipelines using Debezium Change Data Capture (CDC). The project will extend pydbzengine to stream CDC events directly into LangChain and LangGraph by transforming database changes into semantic Document objects and updating vector stores in real time. Key deliverables include: - A Python library integrating Debezium with LangChain/LangGraph - Real-time vector store synchronization (Chroma, PGVector, Milvus) - Pluggable document transformation and ID strategies - End-to-end examples (RAG chatbot, reactive agent, dashboard) - Full documentation, testing, and PyPI release This will provide a production-ready bridge between CDC systems and AI frameworks, ensuring always up-to-date LLM context.
<p>Toil is an open-source Python workflow engine that lets people write data analysis pipelines in Python, CWL, and WDL. Toil has support for common workflow language (CWL), an open standard for describing analysis workflows. The power of Toil was demonstrated in “Toil enables reproducible, open source, big biomedical data analyses” paper published in Nature Biotechnology volume where it is described how well it scaled for a dataset of 108 terabytes on 32,000 cores on a public cloud.</p> <p>This project aims to implement data streaming to speed up the analysis by avoiding slow disk/storage IO and speeding up the start of tool execution when it isn't required to wait for data to download. The main focus is to implement this first in AWS S3.</p>
<p>GenPipes is an important tool by C3G. GenPipes contains a suite of different software. The software contained in the suite is continuously updated but there is no method as of now to automatically update the software catalog which is available on the C3G Website. The goal of the project is to develop a pipeline to automate the updating of the software stack, while also adding some more metadata which would be relevant to the software.</p>