Person v17 was released on 1/10/2022 to Data License Customers
Welcome to our January 2022 release notes! It’s a new year, and we are excited to share all the new updates we’ve been cooking up for you.
We’re kicking things off with a bang this year, and here are some of the key highlights:
- We have officially moved to Monthly API Data updates for our Person Data
- Our new Identify API provides enhanced matching functionality particularly suited for identity risk use cases
- 29 new identity-risk-related fields added to our Person Schema with our new Person Risk attributes
- 30 new fields added to our Company Schema including our new Company Insights Fields targeted toward investment and market research use cases
- We launched a public roadmap for our users to see upcoming/completed product developments as well as suggest and vote for future feature requests.
- Our new AWS Data Exchange API Integration allows users to authenticate API calls using their AWS credentials to centralize data purchasing and processing workflows.
- Over 208 million jobs and 182 million locations updated this quarter.
Millions of new linkages between LinkedIn <> Mobile Phone <> Street Address data
📣 Key Announcements
Deprecations
Schema Changes
This quarter, we are adding 2 new collections of fields to our data:- Person Risk Attributes: A collection of 29 new fields have been added to the Person Schema targeting identity risk and fintech applications.
- Company Insights: A set of 30 new fields that have been added to the Company Schema combining our Person data with company data into aggregated statistics on a company.
New Products and Features
Monthly Data Updates via API
This release also marks our official transition towards supporting monthly data updates for our Person datasets. This has been in the works for a little while now, so here are the details:- Monthly data updates will only be available through API calls, meaning each month the data in our hosted index will be up to date.
- Monthly updates will primarily include data updates and bug fixes.
- New products and field bundles will be rolled out alongside monthly releases on an invite-only basis
- We won’t be rolling out major changes that impact the wider customer base in our monthly releases. As such, we will not be providing monthly release notes.
- For Data License customers, flat file deliveries will continue to be provided on a quarterly basis. We intend to provide new mechanisms for updating flat files like our Retrieve API. As we move into 2022, our goal is to push towards faster and faster updates, which make consuming and delivering flat files difficult. Instead, we invite flat file customers to provide us with feedback on our new mechanisms for data updates.
Identify API
This quarter, we are excited to announce the release of our new Identify API endpoint.This endpoint enables enhanced matching functionality particularly suited for identity risk use cases and includes expanded query parameters and less strict requirements than our Person Enrichment API. It allows you to retrieve multiple strongly associated profiles related to an identity instead of a single best match. This endpoint also includes an improved scoring metric quantifying the matching strength between returned profiles and input parameters. To learn more about the Identify API or to request access please reach to your Customer Success team.Field Bundles
With this release, we are also transitioning to providing curated collections of fields known as Field Bundles as opposed to providing selections of individual fields. These field bundles are designed to be tailored, use-case specific, packages of fields, allowing us to provide value more directly to the problems our users are solving. To support these new field bundles, we have also added over 50 new fields to our Person and Company Data this quarter (see the Person Risk Attributes and Company Insights Fields sections below). Field bundles for person and company data consist of predefined selections of fields from these new additions respectively and complement a common set of base fields that customers universally have access to. As of this release, our Company data is fully bundled, meaning our Company Schema has been transitioned to a set of base fields plus field bundle packages available for purchase. Our Person data is not yet fully bundled, but will be later in 2022.Person Risk Attributes
The Person Risk Attributes field bundle is an addition to the Person Schema consisting of a new set of premium fields designed for identity risk use cases. This field bundle provides new data points such as historically or loosely associated jobs, locations and contact information and enhanced sourcing information for various attributes (like time first/last seen and number of corroborating sources). These fields are provided as a bundle and are integrated into our existing Person-related endpoints (particularly our new Identify API) by providing extended information in the profiles returned.Company Insights Fields
The Company Insights fields are a set of fields added to our Company Schema. These fields were created to help our customers understand company health by looking at the people who make up a company. We used selections of these new Company Insights fields to construct targeted field bundles focusing on specific use cases in the investment and marketing research spaces. Some of the key data points included in this dataset are aggregated month-by-month employee headcounts as well breakdowns of employees by location, role, and seniority. The Company Insights data will be available in multiple bundles to allow you to tap into these fields in a more flexible manner. A highlight that many of our alpha customers have appreciated is the addition of company growth and churn rates, which can be used directly in the Company Search API as filtering criteria.Product Roadmap
We recently introduced a new public-facing product roadmap at feedback.peopledatalabs.com. We wanted to give our customers a way to learn more about what we’re building, and to tell us what they need. Our Product Roadmap is our new interactive platform providing a transparent view of our product development, as well as opportunities to give feedback, submit feature requests, upvote requests, and make your voice heard. For more information on our Product Roadmap, please check out our blog post.AWS Data Exchange API Integration
Our Person Enrichment API has now been officially integrated into the Amazon Data Exchange Marketplace. This new integration allows users to centralize data purchasing and processing workflows and additionally allows users to authenticate their API requests using AWS credentials as an alternative to using their PDL API key. To sign up, you can find the Person Enrichment API listing on the Data Exchange Marketplace.People Data Labs at AWS Re:Invent 2021This integration was also demoed at the AWS Re:Invent 2021 showcase this past quarter, check out the video below!
🚀 Data Updates
Freshness
This quarter, we made huge strides in refreshing our datasets. We updated millions of jobs and locations in our Global Resume Dataset. See below for details:Coverage
Resume Dataset
API Dataset
Mobile Phone Dataset
Email Dataset
Phone Dataset
Street Address Dataset
Commentary
- We decreased the total number of profiles in our API dataset by 28 million, meaning we were able to confidently consolidate millions of fragmented profiles
- We added over 35 million new
experience.location_namesto our resume slice, which is an increase of over 38.8% - We also added over 9 million new roles and 7 million new subroles to experiences in the resume dataset.
- Our resume dataset also saw an increase of over 8 million new
experience.end_dateswhich is a 5.0% improvement in coverage this quarter. - Linkages between Linkedin <> Mobile Phones increased by 8.9%
- We improved our linkages between Street Address <> LinkedIn data by 6.7%
- Mobile Phone <> Street Address linkages increased by 4.3% as well
🛠 Improvements and Bug Fixes
Improvements
- We improved our likelihood scoring for the Person Enrichment API and the new Identify API by building a more accurate probabilistic model that better reflects how records are linked within our datasets.
- We added company filing data into our Company Dataset to improve match rates in our company-related APIs.
- We added a new set of 29 Person Risk Attribute fields to our Person Schema and more than 30 new company fields to our Company Schema including our new Company Insights Fields.
- We improved the error messaging when API users provide invalid inputs containing columns or arrays with more than 100 terms.
- Upgraded our Zapier integration to 1.0.1 which adds a meta-tag for requests so we can better track and support customer requests.
Bug Fixes
- We added a fix to filter out frankenstein records from the Person Enrichment API.
- We fixed a bug with datetime searches in SQL when using our Search APIs.
- We fixed a bug with some profiles in returning non-decoded unicode characters.
- We fixed a bug with capitalization not being accepted by the autocomplete API.
