This survey was conducted in Yemen between March 2013 and July 2014 as part of the joint World Bank/European Bank for Reconstruction and Development (EBRD)/European Investment Bank (EIB) Enterprise Survey. The objective of the survey is to obtain feedback from enterprises on the state of the private sector as well as to help in building a panel of enterprise data that will make it possible to track changes in the business environment over time, thus allowing, for example, impact assessments of reforms. Through interviews with firms in the manufacturing and services sectors, the survey assesses the constraints to private sector growth and creates statistically significant business environment indicators that are comparable across countries.
The standard Enterprise Survey topics include firm characteristics, gender participation, access to finance, annual sales, costs of inputs/labor, workforce composition, bribery, licensing, infrastructure, trade, crime, competition, capacity utilization, land and permits, taxation, informality, business-government relations, innovation and technology, and performance measures. Over 90% of the questions objectively ascertain characteristics of a country's business environment. The remaining questions assess the survey respondents' opinions on what are the obstacles to firm growth and performance.
Kind of data
Sample survey data [ssd]
v02 (August 2015)
- After the exact location of establishments was reviewed, the coding for some observations for variables a3a, a3b, a3c, and a3 was revised. These variables cover location of the establishment and the location size.
- Variables from Manufacturing Questionnaire - Innovation Module and Services Questionnaire - Innovation Module were added to the updated dataset.
v02 of the datasets replaced v01 published in February 2015.
Regions covered are selected based on the number of establishments, contribution to employment, and value added. In most cases these regions are metropolitan areas and reflect the largest centers of economic activity in a country.
Unit of analysis
The primary sampling unit of the study is the establishment. An establishment is a physical location where business is carried out and where industrial operations take place or services are provided. A firm may be composed of one or more establishments. For example, a brewery may have several bottling plants and several establishments for distribution. For the purposes of this survey an establishment must make its own financial decisions and have its own financial statements separate from those of the firm. An establishment must also have its own management and control over its payroll.
The whole population, or universe of the study, is the non-agricultural economy. It comprises: all manufacturing sectors according to the group classification of ISIC Revision 3.1: (group D), construction sector (group F), services sector (groups G and H), and transport, storage, and communications sector (group I). Note that this definition excludes the following sectors: financial intermediation (group J), real estate and renting activities (group K, except sub-sector 72, IT, which was added to the population under study), and all public or utilities-sectors.
Producers and sponsors
European Bank for Reconstruction and Development
European Investment Bank
European Bank for Reconstruction and Development
European Investment Bank
The sample was selected using stratified random sampling. Three levels of stratification were used in this country: industry, establishment size, and region.
Industry stratification was designed in the way that follows: the universe was stratified into one collective manufacturing industry, and two services industries (retail and other services).
Size stratification was defined following the standardized definition for the rollout: small (5 to 19 employees), medium (20 to 99 employees), and large (more than 99 employees). For stratification purposes, the number of employees was defined on the basis of reported permanent full-time workers. This seems to be an appropriate definition of the labor force since seasonal/casual/part-time employment is not common practice, apart from the construction and agriculture sectors which are not included in the survey.
Regional stratification was defined in 6 regions (due to classifications in the sample frame, regions were defined at the governorate level) throughout Yemen. The six regional strata included were Amanat Al-Asemah (Sana'a), Aden, Hudaydah, Hadramaut, Ibb, and Taizz.
For Yemen, two sample frames were used. The first was supplied by the World Bank and consisted of enterprises interviewed in Yemen in 2010. The World Bank required that attempts should be made to re-interview establishments responding to Yemen, Rep. Enterprise Survey 2010, where they met eligibility criteria. This sample is referred to as the panel. The second sample frame, referred to as the fresh sample, was obtained from the Central Statistics Office 2010 Establishment Census, with updates and validation provided by Yemen Polling Center.
The enumerated establishments were then used as the frame for the selection of a sample with the aim of obtaining interviews at 360 establishments with five or more employees. Given the impact that non-eligible units included in the sample universe may have on the results, adjustments may be needed when computing the appropriate weights for individual observations. The percentage of confirmed non-eligible units as a proportion of the total number of sampled establishments contacted for the survey was 14.5% (165 out of 1,141 establishments).
The number of contacted establishments per realized interview was 0.31. This number is the result of two factors: explicit refusals to participate in the survey, as reflected by the rate of rejection (which includes rejections of the screener and the main survey) and the quality of the sample frame, as represented by the presence of ineligible units. The number of rejections per contact was 0.15.
Item non-response was addressed by two strategies:
a- For sensitive questions that may generate negative reactions from the respondent, such as corruption or tax evasion, enumerators were instructed to collect the refusal to respond as a different option from don’t know.
b- Establishments with incomplete information were re-contacted in order to complete this information, whenever necessary.
Survey non-response was addressed by maximizing efforts to contact establishments that were initially selected for interview. Attempts were made to contact the establishment for interview at different times/days of the week before a replacement establishment (with similar strata characteristics) was suggested for interview. Survey non-response did occur but substitutions were made in order to potentially achieve strata-specific goals.
For some units it was impossible to determine eligibility because the contact was not successfully completed. Consequently, different assumptions as to their eligibility result in different universe cells' adjustments and in different sampling weights. Three sets of assumptions were considered:
a- Strict assumption: eligible establishments are only those for which it was possible to directly determine eligibility.
b- Median assumption: eligible establishments are those for which it was possible to directly determine eligibility and those that rejected the screener questionnaire or an answering machine or fax was the only response. Median weights are used for computing indicators on the www.enterprisesurveys.org website.
c- Weak assumption: in addition to the establishments included in points a and b, all establishments for which it was not possible to finalize a contact are assumed eligible. This includes establishments with dead or out of service phone lines, establishments that never answered the phone, and establishments with incorrect addresses for which it was impossible to find a new address. Note that under the weak assumption only observed non-eligible units are excluded from universe projections.
Dates of collection
Mode of data collection
The following survey instruments are available:
- Manufacturing Questionnaire;
- Services Questionnaire.
All variables are named using, first, the letter of each section and, second, the number of the variable within the section, i.e. a1 denotes section A, question 1. Variable names proceeded by a prefix "MNA" indicate questions specific to the Middle East and North Africa region, therefore, they may not be found in the implementation of the rollout in other countries. All other suffixed variables are global and are present in all economy surveys over the world. All variables are numeric with the exception of those variables with an "x" at the end of their names. The suffix "x" denotes that the variable is alpha-numeric.
There are two establishment identifiers, idstd and id. The first is a global unique identifier. The second is a country unique identifier. The variables a2 (sampling region), a6a (sampling establishment's size), and a4a (sampling sector) contain the establishment's classification into the strata chosen for each country using information from the sample frame. The strata were defined according to the guidelines described above. Variable a4a is coded using ISIC Rev 3.1 codes for the chosen industries for stratification. These codes include most manufacturing industries (15 to 37), retail (52), and (45, 50, 51, 55, 60-64, 72) for other services.
Yemen Polling Center
Data entry and quality controls are implemented by the contractor and data is delivered to the World Bank in batches (typically 10%, 50% and 100%). These data deliveries are checked for logical consistency, out of range values, skip patterns, and duplicate entries. Problems are flagged by the World Bank and corrected by the implementing contractor through data checks, callbacks, and revisiting establishments.
Confidentiality of the survey respondents and the sensitive information they provide is necessary to ensure the greatest degree of survey participation, integrity and confidence in the quality of the data. Surveys are usually carried out in cooperation with business organizations and government agencies promoting job creation and economic growth, but confidentiality is never compromised.
The use of the datasets must be acknowledged using a citation which would include:
- the identification of the Primary Investigator (including country name);
- the full title of the survey and its acronym (when available), and the year(s) of implementation;
- the survey reference number;
- the source and date of download (for datasets disseminated online).
World Bank, European Bank for Reconstruction and Development, European Investment Bank. Yemen, Rep. Enterprise Survey (ES) 2013, Ref. YEM_2013_ES_v02_M. Dataset downloaded from [URL] on [date].
Disclaimer and copyrights
The user of the data acknowledges that the original collector of the data, the authorized distributor of the data, and the relevant funding agency bear no responsibility for use of the data or for interpretations or inferences based upon such uses.
Enterprise Analysis Unit
Development Data Group
v02 (December 2015)
- The dataset was updated
- Manufacturing Questionnaire - Innovation Module and Services Questionnaire - Innovation Module were added to Related Materials