Federal grant · project grant (b)
I-corps: Translation Potential of Synthetic Data Generation With Nullspace Sampling for Tabular and Timeseries Data -the Broader Impact of This I-corps Project Is Based on the Development of Software to Generate Synthetic Data for Use in the Healthcare, Consulting, and Insurance Industries. Synthetic Data Is Artificially Generated Data That Is Statistically Similar to Real-world Datasets Used by Businesses. Synthetic Data Can Be Used for Analytics and Machine Learning When Access to Real Data Is Limited and May Have Uses in Augmenting Minority Representation in Real-world Datasets Thereby Aiding in More Equitable Outcomes. Overall, the Broad Applicability of Synthetic Datasets Has the Potential to Drive Innovation in Healthcare and Other Industries by Allowing Businesses to Share Synthetic Versions of Proprietary Data With Strategic Partners, Such as Data Analytics Companies, and Remain in Full Compliance With Data Privacy Laws. This Ability Can Lead to an Increase in Data-driven Decision-making in the Private Sector and Effective Policy Formulation in the Public Sector. for Instance, Applications of Synthetic Medical Data May Help Healthcare Researchers and Administrators to Better Model Patient Activity, Including Representative Data of Understudied Populations, and Ultimately Improve Human Health. This I-corps Project Utilizes Experiential Learning Coupled With a First-hand Investigation of the Industry Ecosystem to Assess the Translation Potential of the Technology. the Solution Is Based on the Prior Development of a Non-deep Learning Technique to Generate Synthetic Datasets Using Features of Real Data. Synthetic Data Is Artificially Generated Data That Is Statistically Similar to Real Datasets and Can Be Used for Analytics and Machine Learning When Access to Real Data Is Limited. This Innovative Solution Allows Significantly Faster Generation of Tabular and Timeseries Synthetic Data Without the Need for Training or Optimization Processes, While Internally Using Linear Algebra-based Techniques. Although This Solution Was Initially Created to Generate Synthetic Timeseries Data, It Can Be Modified to Generate Synthetic Tabular Data. This Solution Is Completely Non-parametric and Does Not Involve the Additional Steps Associated With Training and Optimization, Making It 300X Faster Than State-of-the-art Deep Learning Generation Methods for Tabular Data. Thus, This Approach Can Generate Richly Structured Datasets Using Significantly Less Computing Time Relative to Deep-learning Methods. This Award Reflects NSF'S Statutory Mission and Has Been Deemed Worthy of Support Through Evaluation Using the Foundation's Intellectual Merit and Broader Impacts Review Criteria.- Subawards Are Not Planned for This Award.
Committed
$50,000
Loading…
Everything here is this single award's whole record — signed, amended, paid — not a fiscal-year slice. The by-year charts elsewhere split an award across the years it was committed; this page keeps it whole.
Committed is what the government has legally promised on this award so far. Contracts can also carry a ceiling — the maximum if every option is exercised. Unspent ceiling is headroom, not money owed.
The cash actually disbursed against this award. The gap from committed is the disbursement pipeline: promised, not yet cashed.
Each transaction is a signing event — an action that created or changed the award, dated the day it was signed — not a payment. Negative amounts are real: money de-committed at closeout or renegotiation.
One bar, the award’s whole arithmetic: paid out, then committed, not yet paid, then unspent ceiling.